spare
spareqwen3-30b-a3b-0703-fixed-rlve-official16-iter351qwen3-30b-a3b-0703-fixed-rlve-official16-iter367qwen3-30b-a3b-0703-fixed-rlve-official16-iter383qwen3-30b-a3b-0703-fixed-rlve-official16-iter319nemotron-3-nano-30b-20260719-games-iter39Qwen2.5-3b-spare-prm-mathqwen3-30b-a3b-0703-fixed-rlve-official16-iter399
Datasets
All datasets matching “spare”nemotron-3-nano-30b-20260719-spare-games-envs
Nemotron-3-Nano-30B SPARE Self-Play Environments (run_20260719_final)
This dataset packages the self-play generated game environments produced
by a live SPARE (Self-Play with Adaptive cuRriculum Extension) training run
of NVIDIA-Nemotron-3-Nano-30B-A3B. It is a raw-data export for another
agent to pick up, replay, and build its own visualization / weave log from.
Provenance
Run: run_20260719_final
Source Ray job: spare_nemotron_games_mtpg768_1784556397 (the live… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/nemotron-3-nano-30b-20260719-spare-games-envs.spare-fixed-corpus-gpt53-6skill
SPARE fixed corpus — gpt-5.3, 6 skills
Balanced 2,400-game validated corpus (400 per skill) generated by gpt-5.3-chat, the external-generator counterpart to the self-generated 30B corpus. Top-level games are the kept set; _unused/ and _invalid/ retain the full validation record.
2400 validated environments.
Skill
Games
Causal Inference
400
Logical Deduction
400
Mathematical Reasoning
400
Optimization
400
Pattern Recognition
400
Spatial Reasoning
400… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/spare-fixed-corpus-gpt53-6skill.qwen3-30b-plateau-kl0-spare-games-envs
qwen3-30B-A3B-Instruct plateau-6skill KL=0 — generated environments
Environments generated during the 30B plateau KL=0 run (2026-07-19), recovered from the run's surviving on-disk game cache.
Games
640
Steps covered
46 (step 0–96)
Skill
Games
Causal Inference
96
Logical Deduction
115
Mathematical Reasoning
120
Optimization
97
Pattern Recognition
100
Spatial Reasoning
112
Layout
manifest.json… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-30b-plateau-kl0-spare-games-envs.qwen3-30b-plateau-kl005-spare-games-envs
qwen3-30B-A3B-Instruct plateau-6skill KL=0.05 — generated environments
Environments generated during the 30B plateau KL=0.05 run (main segment 20260718_103657), recovered from the run's surviving on-disk game cache.
Games
480
Steps covered
24 (step 0–76)
Skill
Games
Causal Inference
81
Logical Deduction
80
Mathematical Reasoning
81
Optimization
79
Pattern Recognition
80
Spatial Reasoning
79
Layout
manifest.json… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-30b-plateau-kl005-spare-games-envs.qwen3-30b-0705c-glory-r8-premerge-spare-games-envs
qwen3-30B-A3B-Instruct-0705c-glory-r8-premerge — generated environments
Environments generated by the SPARE proposer during training run
5hvg1dna (qwen3-30B-A3B-Instruct-0705c-glory-r8-premerge), recovered from the spare-viz durable cache.
The run's scratch directory no longer exists; this dataset is the surviving copy.
Games
4023
Steps covered
105 (step 0–392)
With recovered skill
4023
With hint
0
Actor / proposer model… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-30b-0705c-glory-r8-premerge-spare-games-envs.spare-gpt55-static-corpus
SPARE GPT-5.5 Grounded Cognitive Multi-Turn Games
This public dataset contains 7,872 validated Python game environments for actor-only SPARE training.
Six cognitive skills, exactly 1,312 environments per skill
Generated with GPT-5.5 and grounded by spice_megascience_15k.jsonl
Grounding corpus SHA-256: a36a928b4940b5b5d9e3f4cb5804a94c69462360943adb3be14613c82f0f72c0
Maximum 25 turns and 32K generation context
Every environment passes load, reset, step, and replay validation with… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/spare-gpt55-static-corpus.
