CoolFace
Datasetpublic

FineEnvs/SmolDataEnvs-harbor-train

📊 SmolDataEnvs: Harbor (train) 5.5K+ RL tasks for hill-climbing small models in code and data science. A 2B model on these tasks. Left: what it optimises. Right: 144 held-out tasks it never trains on. Two runs over the same 5,000 tasks: shuffled against a curriculum ordered easiest to hardest. The training suite: 5,000 hands-on data-analysis tasks. Each one drops an agent into a sandbox with a real dataset and a question, and asks it to explore the data… See the full description on the dataset page: https://huggingface.co/datasets/FineEnvs/SmolDataEnvs-harbor-train.

sourceHugging Facemitupdated 21h agoView on Hugging Face
1likes73downloads
filemanifest.parquet125 KBdownload

FineEnvs/SmolDataEnvs-harbor-train · main · files are served by the source, never re-hosted here

FineEnvs/SmolDataEnvs-harbor-train · CoolFace