CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01CoopReason /TESSY-Code-80K TESSY-Code-80K 📄 Paper Link    |    🔗 GitHub Repository 📣 Paper 🎉 Accepted at ICML 2026! How to Fine-Tune a Reasoning Model? A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data 🚀 Overview We construct a programming contest training dataset for Qwen3-8B by leveraging GPT-OSS-120B as the teacher model. The synthesized data preserves the strong reasoning capabilities of GPT-OSS-120B, while being aligned with the… See the full description on the dataset page: https://huggingface.co/datasets/CoopReason/TESSY-Code-80K.texttext-generation10K<n<100K10 likes632 downloads5mo agoHugging Face02CoopReason /TESSY-SuperGPQA-3K TESSY-SuperGPQA-3K 📄 Paper Link    |    🔗 GitHub Repository 📣 Paper 🎉 Accepted at ICML 2026! How to Fine-Tune a Reasoning Model? A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data 🚀 Overview We construct a programming contest training dataset for Qwen3-8B by leveraging GPT-OSS-120B as the teacher model. The synthesized data preserves the strong reasoning capabilities of GPT-OSS-120B, while being aligned with the… See the full description on the dataset page: https://huggingface.co/datasets/CoopReason/TESSY-SuperGPQA-3K.texttext-generation1K<n<10K6 likes162 downloads5mo agoHugging Face03CoopReason /TESSY-Math-12K TESSY-Math-12K 📄 Paper Link    |    🔗 GitHub Repository 📣 Paper 🎉 Accepted at ICML 2026! How to Fine-Tune a Reasoning Model? A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data 🚀 Overview We construct a programming contest training dataset for Qwen3-8B by leveraging GPT-OSS-120B as the teacher model. The synthesized data preserves the strong reasoning capabilities of GPT-OSS-120B, while being aligned with the… See the full description on the dataset page: https://huggingface.co/datasets/CoopReason/TESSY-Math-12K.texttext-generation10K<n<100K6 likes140 downloads5mo agoHugging Face04iamPi /tessera-8a92b237 tessera — corpus epoch 13, full sweep Teacher-anchored SFT data harvested from every published Affine (Bittensor SN120) duel scored against corpus epoch 13 — 230 duel records, chal-00760 through chal-01102, covering 2026-08-16 to 2026-08-24. 42,006 rows over 42,006 distinct turns (one row per turn), drawn from 4,981 trajectories and 3,751 strata. That is 70% of the 59,745-turn epoch-13 corpus, and 2.3× the 18,138 rows of iamPi/tessera-77d11909, which sampled a subset of the same… See the full description on the dataset page: https://huggingface.co/datasets/iamPi/tessera-8a92b237.tabulartext-generation10K<n<100K0 likes102 downloads1mo agoHugging Face05smirki /Agentic-Coding-Tessa Agentic Coding Dataset for Tessa A comprehensive dataset for training coding agents with tool-use, reasoning, and software engineering capabilities. Dataset Composition This dataset combines multiple high-quality sources: hermes_reasoning (20.0%): Tool-use and reasoning dataset - interstellarninja/hermes_reasoning_tool_use search_arena (15.0%): Search and retrieval tasks - lmarena-ai/search-arena-24k arena_human_pref (15.0%): Human preference data for alignment -… See the full description on the dataset page: https://huggingface.co/datasets/smirki/Agentic-Coding-Tessa.texttext-generation10K<n<100K13 likes51 downloads1y agoHugging Face06Tesslate /Gradient-Reasoningtexttext-generation10K<n<100K7 likes36 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.