CoolFace
Datasetpublic

FineEnvs/SmolDataEnvs-harbor-eval

📊 SmolDataEnvs: Harbor (eval) 5.5K+ RL tasks for hill-climbing small models in code and data science. A 2B model on these tasks. Left: what it optimises. Right: 144 held-out tasks it never trains on. Two runs over the same 5,000 tasks: shuffled against a curriculum ordered easiest to hardest. The validation suite: 144 tasks, small enough to run every few hundred training steps without the eval becoming the expensive part of the loop. Packaged in Harbor format… See the full description on the dataset page: https://huggingface.co/datasets/FineEnvs/SmolDataEnvs-harbor-eval.

sourceHugging Facemitupdated 8h agoView on Hugging Face
0likes
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face