datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
prompt-swap-mixed12-5xlr-e1-mxfp4-mergedprompt-swap-mixed12-5xlr-e2-mxfp4-mergedprompt-swap-medium12-e2-mxfp4-mergedai-models-2026
AI Models & Releases 2026
AI model releases, benchmarks, capabilities. Updated daily via automated collection pipeline.
Part of the Legion Data Factory — historical AI ecosystem datasets 2026.
Methodology
Automated collection from public sources (HackerNews, RSS feeds, APIs).
Updated daily via cron job. Raw data, minimal processing.
License
CC BY 4.0
🔑 API Access — Updated Daily
Live data via Legion AI API | Documentation
Free: 100… See the full description on the dataset page: https://huggingface.co/datasets/gemmozero/ai-models-2026.aimo3-train-dataaimo3-math-dataset
AIMO3 Math Dataset
Training data for AI Mathematical Olympiad Progress Prize 3.
Files
train_cot.jsonl - Chain-of-Thought examples
train_tir.jsonl - Tool-Integrated Reasoning examples
Author
Ryan J Cardwell (Archer Phoenix) - AIMO3 Competitor
aimodel-sft-v1Train_dataAI-MO-NuminaMath-TIR-korean-240918
IMPORTANT NOTE
This data is part of the progress. Current translation progress: 24.85% (2024-09-18 01:32 KST)
I'm taking a short break due to personal reasons. I'll be back in a month.
TODO-LIST
Finish translation
Translation
I used gemini-1.5-pro-exp-0827. The prompt used for translation will be disclosed at the end.
Dataset Card for NuminaMath CoT
Dataset Summary
Tool-integrated reasoning (TIR) plays a crucial role in this… See the full description on the dataset page: https://huggingface.co/datasets/ChuGyouk/AI-MO-NuminaMath-TIR-korean-240918.GeometryLeanBenchprompt-swap-medium12-e1-mxfp4-mergedai-model-deprecation-and-retirement
AI model deprecation and retirement dates by provider
Canonical, always-current version: https://referencesource.org/ai-model-deprecation-and-retirement/
Machine-readable: https://referencesource.org/ai-model-deprecation-and-retirement/data.json — this mirror is a point-in-time copy.
Last verified: 2026-08-12
Stale after: 2026-09-11 (past this date, prefer the canonical copy —
it re-verifies on a cadence this snapshot does not)
Records: 347
Which AI API models are deprecated… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/ai-model-deprecation-and-retirement.AI-MO__NuminaMath-7B-TIR-details
Dataset Card for Evaluation run of AI-MO/NuminaMath-7B-TIR
Dataset automatically created during the evaluation run of model AI-MO/NuminaMath-7B-TIR
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/AI-MO__NuminaMath-7B-TIR-details.AI-MO__NuminaMath-7B-CoT-details
Dataset Card for Evaluation run of AI-MO/NuminaMath-7B-CoT
Dataset automatically created during the evaluation run of model AI-MO/NuminaMath-7B-CoT
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/AI-MO__NuminaMath-7B-CoT-details.dataframer_text_to_sqltraining-data-oss120bAIMO3_to_Hardbrian-rollouts-311-training-turns
brian-rollouts-311-training-turns-v1
Lean per-turn GPT-OSS training export derived from
brian-rollouts-311-full-v2.
Contents:
rollouts.training_turns.jsonl: one row per assistant call with full prompt token IDs and completion token IDs
manifest.json: export metadata and row counts
Semantics:
prompt_token_ids are the full context shown to the model for that assistant call.
completion_token_ids are the tokens generated by the model on that call.
prompt_token_ids include system… See the full description on the dataset page: https://huggingface.co/datasets/aimosprite/brian-rollouts-311-training-turns.ai-model-accent-corpus
AI Model Accent Corpus
A paired corpus for studying the prose "accent" of large language models: 5 models × the same 102 open-ended prompts = 510 plain-prose passages, generated June 2026 at a fixed decoding temperature, constrained to plain paragraphs (no lists or headings) so the data isolates prose style rather than formatting choices.
Models: OpenAI GPT-4o, GPT-4o-mini, GPT-3.5-turbo; Anthropic Claude Sonnet 4.5, Claude Haiku 4.5.
Released with the study "Every Model Has an… See the full description on the dataset page: https://huggingface.co/datasets/firatmihci/ai-model-accent-corpus.kaggle-aimo2aimo-validation-amc-repeated3moh_8_fake_rollouts
MOH-8 Fake Rollouts
480 math competition problems, each with 8 candidate solution rollouts from OSS 120B.
A controlled number of rollouts per problem are correct — use this to train/test a
verifier model that must identify which solutions are right.
Source
Problems and rollouts sampled from aimosprite/training-data-oss120b (the oss128-fixed-FINAL.jsonl file).
Only polymath-source problems in the 2/8–4/8 pass rate range (32–64 correct out of 128 attempts).
4 problems… See the full description on the dataset page: https://huggingface.co/datasets/aimosprite/moh_8_fake_rollouts.Test_DataAIMO-2_ReferenceThis CSV file is reference.csv in Kaggle's AI Mathematical Olympiad - Progress Prize 2.
training-settest-largedate:
mar 02
gpt-oss-120b-marina-cleaned-tok-idchinese-translatedbrian-benchaimo-phase1-teacher-corpus
AIMO Phase 1 Teacher Corpus
Private export of the current Phase 1 teacher-style corpus used for overnight data preparation.
source file: runtime-teacher-phase1.jsonl
