CoolFace
Datasetpublic

FineEnvs/SmolDataEnvs-harbor-eval

📊 SmolDataEnvs: Harbor (eval) 5.5K+ RL tasks for hill-climbing small models in code and data science. A 2B model on these tasks. Left: what it optimises. Right: 144 held-out tasks it never trains on. Two runs over the same 5,000 tasks: shuffled against a curriculum ordered easiest to hardest. The validation suite: 144 tasks, small enough to run every few hundred training steps without the eval becoming the expensive part of the loop. Packaged in Harbor format… See the full description on the dataset page: https://huggingface.co/datasets/FineEnvs/SmolDataEnvs-harbor-eval.

sourceHugging Facemitupdated 17h agoView on Hugging Face
0likes132downloads

Nothing at this path on main. The folder may be empty, or the revision may not exist.

FineEnvs/SmolDataEnvs-harbor-eval · main · files are served by the source, never re-hosted here