datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
toy-models-of-sft-data
Toy Models of SFT Data
This is a public-clean candidate data package for the Toy Models of SFT project.
It is built for researcher inspection first.
The package answers two questions:
What were the models trained on?
How did the models actually behave under evaluation?
The package includes training data, eval inputs, model rollouts, judge scores,
parsed GPQA outputs, aggregate tables, paper figures, frozen plot data, and
provenance records. It deliberately includes some… See the full description on the dataset page: https://huggingface.co/datasets/matonski/toy-models-of-sft-data.coderTraining dataset for finetuning for human-eval.
This dataset has been created from the following datasets:
sahil2801/CodeAlpaca-20k
sahil2801/code_instructions_120k
mhhmm/leetcode-solutions-python
teknium1/GPTeacher
Script for generating dataset: create_dataset.py.
