CoolFace
20 results

spare

msr-spare-1 /nemotron-3-nano-30b-20260719-spare-games-envs Nemotron-3-Nano-30B SPARE Self-Play Environments (run_20260719_final) This dataset packages the self-play generated game environments produced by a live SPARE (Self-Play with Adaptive cuRriculum Extension) training run of NVIDIA-Nemotron-3-Nano-30B-A3B. It is a raw-data export for another agent to pick up, replay, and build its own visualization / weave log from. Provenance Run: run_20260719_final Source Ray job: spare_nemotron_games_mtpg768_1784556397 (the live… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/nemotron-3-nano-30b-20260719-spare-games-envs.textn<1K0 likes6.1k downloads2mo agoHugging Facemsr-spare-1 /spare-fixed-corpus-gpt53-6skill SPARE fixed corpus — gpt-5.3, 6 skills Balanced 2,400-game validated corpus (400 per skill) generated by gpt-5.3-chat, the external-generator counterpart to the self-generated 30B corpus. Top-level games are the kept set; _unused/ and _invalid/ retain the full validation record. 2400 validated environments. Skill Games Causal Inference 400 Logical Deduction 400 Mathematical Reasoning 400 Optimization 400 Pattern Recognition 400 Spatial Reasoning 400… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/spare-fixed-corpus-gpt53-6skill.0 likes1k downloads1mo agoHugging Facemsr-spare-1 /qwen3-30b-plateau-kl0-spare-games-envs qwen3-30B-A3B-Instruct plateau-6skill KL=0 — generated environments Environments generated during the 30B plateau KL=0 run (2026-07-19), recovered from the run's surviving on-disk game cache. Games 640 Steps covered 46 (step 0–96) Skill Games Causal Inference 96 Logical Deduction 115 Mathematical Reasoning 120 Optimization 97 Pattern Recognition 100 Spatial Reasoning 112 Layout manifest.json… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-30b-plateau-kl0-spare-games-envs.textn<1K0 likes907 downloads1mo agoHugging Facemsr-spare-1 /qwen3-30b-plateau-kl005-spare-games-envs qwen3-30B-A3B-Instruct plateau-6skill KL=0.05 — generated environments Environments generated during the 30B plateau KL=0.05 run (main segment 20260718_103657), recovered from the run's surviving on-disk game cache. Games 480 Steps covered 24 (step 0–76) Skill Games Causal Inference 81 Logical Deduction 80 Mathematical Reasoning 81 Optimization 79 Pattern Recognition 80 Spatial Reasoning 79 Layout manifest.json… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-30b-plateau-kl005-spare-games-envs.textn<1K0 likes861 downloads1mo agoHugging Facemsr-spare-1 /qwen3-30b-0705c-glory-r8-premerge-spare-games-envs qwen3-30B-A3B-Instruct-0705c-glory-r8-premerge — generated environments Environments generated by the SPARE proposer during training run 5hvg1dna (qwen3-30B-A3B-Instruct-0705c-glory-r8-premerge), recovered from the spare-viz durable cache. The run's scratch directory no longer exists; this dataset is the surviving copy. Games 4023 Steps covered 105 (step 0–392) With recovered skill 4023 With hint 0 Actor / proposer model… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/qwen3-30b-0705c-glory-r8-premerge-spare-games-envs.0 likes672 downloads1mo agoHugging Facemsr-spare-1 /spare-gpt55-static-corpus SPARE GPT-5.5 Grounded Cognitive Multi-Turn Games This public dataset contains 7,872 validated Python game environments for actor-only SPARE training. Six cognitive skills, exactly 1,312 environments per skill Generated with GPT-5.5 and grounded by spice_megascience_15k.jsonl Grounding corpus SHA-256: a36a928b4940b5b5d9e3f4cb5804a94c69462360943adb3be14613c82f0f72c0 Maximum 25 turns and 32K generation context Every environment passes load, reset, step, and replay validation with… See the full description on the dataset page: https://huggingface.co/datasets/msr-spare-1/spare-gpt55-static-corpus.textreinforcement-learning1K<n<10K0 likes592 downloads2mo agoHugging Face