CoolFace
Datasetpublic

novastar111/sokoban_easy_cot_chunk_kinf_train

sokoban_easy_cot_chunk_kinf_train BAGEL VLM-Gym world-model dataset (sokoban / cot). CoT chunk-K train set: all-step interleaved imagined reasoning; re-grounds on the true frame every K=inf (open-loop; imagine the whole episode) steps. layout: Train-only. Gzipped-JSONL shards under training/; each row is one packed SFT sample with base64-JPEG frames inline. images are base64-encoded JPEG frames stored inline in each JSONL row. Pairs with the matching sokoban checkpoint(s)… See the full description on the dataset page: https://huggingface.co/datasets/novastar111/sokoban_easy_cot_chunk_kinf_train.

sourceHugging Facecc-by-nc-4.0updated 1mo agoView on Hugging Face
0likes15downloads
Dataset Card

sokobaneasycotchunkkinf_train

BAGEL VLM-Gym world-model dataset (sokoban / cot).

  • —CoT chunk-K train set: all-step interleaved imagined reasoning; re-grounds on the true frame every K=inf (open-loop; imagine the whole episode) steps.
  • —layout: Train-only. Gzipped-JSONL shards under training/; each row is one packed SFT sample with base64-JPEG frames inline.
  • —images are base64-encoded JPEG frames stored inline in each JSONL row.

Pairs with the matching sokoban checkpoint(s) under the companion model org; CoT and non-CoT variants share the same underlying rollouts.