FineEnvs/SmolDataEnvs-harbor-test
📊 SmolDataEnvs: Harbor (test) 5.5K+ RL tasks for hill-climbing small models in code and data science. A 2B model on these tasks. Left: what it optimises. Right: 144 held-out tasks it never trains on. Two runs over the same 5,000 tasks: shuffled against a curriculum ordered easiest to hardest. The held-out benchmark: 250 tasks the model never trains on, weighted towards the hard end on purpose. This is the suite to quote a number from. Packaged in Harbor format… See the full description on the dataset page: https://huggingface.co/datasets/FineEnvs/SmolDataEnvs-harbor-test.
Card: drop em dashes
Card: FineEnvs citation, install line for openenv[harbor]
Card: full banner on top
Card GIF: hill climb only, no banner band
README: the curves loop now carries the banner
README: light-theme curves, split train | eval, shuffled vs curriculum
README: quote the tagline, add the training-curve loop
SmolDataEnvs: rename, new README, banner
docs: link the Harbor Visualiser from the top of the card
tags: declare as an RL environment dataset
grader: stop emitting tool_efficiency, which was always null and destroyed correctness
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
initial commit
