novastar111/maze2d_easy_noncot_chunk_k1_train
maze2d_easy_noncot_chunk_k1_train BAGEL VLM-Gym world-model dataset (maze2d / noncot). Non-CoT chunk-K train set (no imagined reasoning); re-grounds every K=1 steps. layout: Train-only. Gzipped-JSONL shards under training/; each row is one packed SFT sample with base64-JPEG frames inline. images are base64-encoded JPEG frames stored inline in each JSONL row. Pairs with the matching maze2d checkpoint(s) under the companion model org; CoT and non-CoT variants share the same… See the full description on the dataset page: https://huggingface.co/datasets/novastar111/maze2d_easy_noncot_chunk_k1_train.
maze2deasynoncotchunkk1_train
BAGEL VLM-Gym world-model dataset (maze2d / noncot).
- Non-CoT chunk-K train set (no imagined reasoning); re-grounds every K=1 steps.
- layout: Train-only. Gzipped-JSONL shards under
training/; each row is one packed SFT sample with base64-JPEG frames inline. - images are base64-encoded JPEG frames stored inline in each JSONL row.
Pairs with the matching maze2d checkpoint(s) under the companion model org; CoT and non-CoT variants share the same underlying rollouts.
