datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SmolDataEnvs
📈 SmolDataEnvs
5.5K+ RL tasks for hill-climbing small models in code and data science.
A 2B model on these tasks. Left: what it optimises. Right: 144 held-out tasks it never trains on.
Two runs over the same 5,000 tasks: shuffled against a curriculum ordered easiest to hardest.
Data-analysis tasks as a plain, load-and-go dataset: no runtime, no framework required. Each
row is one self-contained task: a real tabular dataset, a question about it, and a gold answer… See the full description on the dataset page: https://huggingface.co/datasets/FineEnvs/SmolDataEnvs.SmolDataEnvs-sft
🛠️ SmolDataEnvs: SFT
5.5K+ RL tasks for hill-climbing small models in code and data science.
A 2B model on these tasks. Left: what it optimises. Right: 144 held-out tasks it never trains on.
Two runs over the same 5,000 tasks: shuffled against a curriculum ordered easiest to hardest.
4,677 worked examples of an agent doing data science the right way. Each row is a complete,
verified-correct trajectory: read the question, poke at the data with a shell tool… See the full description on the dataset page: https://huggingface.co/datasets/FineEnvs/SmolDataEnvs-sft.
