FineEnvs/SmolDataEnvs-harbor-eval
📊 SmolDataEnvs: Harbor (eval) 5.5K+ RL tasks for hill-climbing small models in code and data science. A 2B model on these tasks. Left: what it optimises. Right: 144 held-out tasks it never trains on. Two runs over the same 5,000 tasks: shuffled against a curriculum ordered easiest to hardest. The validation suite: 144 tasks, small enough to run every few hundred training steps without the eval becoming the expensive part of the loop. Packaged in Harbor format… See the full description on the dataset page: https://huggingface.co/datasets/FineEnvs/SmolDataEnvs-harbor-eval.
This repository belongs to FineEnvs on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
