datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
TESSY-Code-80K
TESSY-Code-80K
📄 Paper Link
|
🔗 GitHub Repository
📣 Paper
🎉 Accepted at ICML 2026!
How to Fine-Tune a Reasoning Model? A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data
🚀 Overview
We construct a programming contest training dataset for Qwen3-8B by leveraging GPT-OSS-120B as the teacher model. The synthesized data preserves the strong reasoning capabilities of GPT-OSS-120B, while being aligned with the… See the full description on the dataset page: https://huggingface.co/datasets/CoopReason/TESSY-Code-80K.TESSY-SuperGPQA-3K
TESSY-SuperGPQA-3K
📄 Paper Link
|
🔗 GitHub Repository
📣 Paper
🎉 Accepted at ICML 2026!
How to Fine-Tune a Reasoning Model? A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data
🚀 Overview
We construct a programming contest training dataset for Qwen3-8B by leveraging GPT-OSS-120B as the teacher model. The synthesized data preserves the strong reasoning capabilities of GPT-OSS-120B, while being aligned with the… See the full description on the dataset page: https://huggingface.co/datasets/CoopReason/TESSY-SuperGPQA-3K.TESSY-Math-12K
TESSY-Math-12K
📄 Paper Link
|
🔗 GitHub Repository
📣 Paper
🎉 Accepted at ICML 2026!
How to Fine-Tune a Reasoning Model? A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data
🚀 Overview
We construct a programming contest training dataset for Qwen3-8B by leveraging GPT-OSS-120B as the teacher model. The synthesized data preserves the strong reasoning capabilities of GPT-OSS-120B, while being aligned with the… See the full description on the dataset page: https://huggingface.co/datasets/CoopReason/TESSY-Math-12K.tessera-8a92b237
tessera — corpus epoch 13, full sweep
Teacher-anchored SFT data harvested from every published Affine (Bittensor
SN120) duel scored against corpus epoch 13 — 230 duel records, chal-00760
through chal-01102, covering 2026-08-16 to 2026-08-24.
42,006 rows over 42,006 distinct turns (one row per turn), drawn from
4,981 trajectories and 3,751 strata. That is 70% of the 59,745-turn epoch-13
corpus, and 2.3× the 18,138 rows of
iamPi/tessera-77d11909,
which sampled a subset of the same… See the full description on the dataset page: https://huggingface.co/datasets/iamPi/tessera-8a92b237.Agentic-Coding-Tessa
Agentic Coding Dataset for Tessa
A comprehensive dataset for training coding agents with tool-use, reasoning, and software engineering capabilities.
Dataset Composition
This dataset combines multiple high-quality sources:
hermes_reasoning (20.0%): Tool-use and reasoning dataset - interstellarninja/hermes_reasoning_tool_use
search_arena (15.0%): Search and retrieval tasks - lmarena-ai/search-arena-24k
arena_human_pref (15.0%): Human preference data for alignment -… See the full description on the dataset page: https://huggingface.co/datasets/smirki/Agentic-Coding-Tessa.Gradient-Reasoning
