datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
small-magpie
Smaller Magpie
A collection of smaller Magpie datasets compared to agentlans/magpie.
For argilla/magpie-ultra-v0.1, only instructions rated as good or excellent were selected.
output_quality corresponds to the original dataset’s score_difference, which is the gap between instruct model and base model responses as evaluated by a reward model.
Please see the original dataset for details.
Source
Rows
argilla/magpie-ultra-v0.1
43923
Mxode/Magpie-Pro-10K-GPT4o-mini10000
small-mind-pmb-v0
PMB v0 — Personalised Memory Benchmark
An evaluation benchmark for long-horizon personalised memory in small language models. It asks
whether a model can recall what a specific user told it across many sessions, and — the part most
memory benchmarks skip — whether it can decline to answer when the memory does not contain the
answer.
Built for small-mind-companion, a study of how much of the long-horizon memory gap a
~2B multimodal model can close without scaling parameters.
Part… See the full description on the dataset page: https://huggingface.co/datasets/arjhinety/small-mind-pmb-v0.small-model-schema-gym
Small Model Schema Gym Dataset
Deterministically generated chat examples for first-pass JSON compliance on
the project Dream brief and Safety plan contracts.
Files
train.jsonl: 500 training examples.
validation.jsonl: 200 held-out examples.
manifest.json: counts, seed, provenance, and overlap check.
Each row contains:
{
"id": "stable example identifier",
"messages": [
{"role": "system", "content": "..."},
{"role": "user", "content": "..."}… See the full description on the dataset page: https://huggingface.co/datasets/KwabsHug/small-model-schema-gym.small_molecule_drugsPileV2-axolotlformat-smallmixsmall-moviessmall_model_datasetsmall-model-simulation-data
