CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01beatsprom /deepseek-r1-systems-kernel-reasoning 🧠 DeepSeek-R1 Low-Level Systems & Kernel Reasoning Suite (2026) 🛒 Commercial Full Suite Available: The full production suite with 10,000 SFT Hardware Reasoning Traces + 2,500 High-Contrast DPO Alignment Pairs across all 20 domains is available on Gumroad: 👉 Download Full Commercial Dataset on Gumroad (Starter \ / Pro \ / Enterprise ) A Tier-1 Commercial Dataset Suite engineered specifically for fine-tuning DeepSeek-R1, DeepSeek-R1-Distill-Qwen-14B/32B, and frontier… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/deepseek-r1-systems-kernel-reasoning.tabulartext-generation1K<n<10K0 likes46 downloads13d agoHugging Face02jrosseruk /DeepSeek-R1-Distill-Llama-8B-MATH-traces DeepSeek-R1-Distill-Llama-8B MATH Reasoning Traces 10,000 reasoning traces from DeepSeek-R1-Distill-Llama-8B on MATH problems. Model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B (served via vLLM) Source problems: xDAN2099/lighteval-MATH (train split) Sampling: 500 problems (100 per difficulty level 1-5) x 20 rollouts Generation params: temperature=0.6, top_p=0.95, max_tokens=15000 Accuracy: 80.1% (8,008 correct / 1,992 incorrect) Problem types: Algebra, Counting & Probability… See the full description on the dataset page: https://huggingface.co/datasets/jrosseruk/DeepSeek-R1-Distill-Llama-8B-MATH-traces.tabulartext-generation10K<n<100K0 likes23 downloads8mo agoHugging Face03jrosseruk /DeepSeek-R1-Distill-Llama-8B-MATH-labeled-sentences DeepSeek-R1-Distill-Llama-8B MATH Labeled Sentences Sentence-level function-tag labels for reasoning traces from DeepSeek-R1-Distill-Llama-8B on MATH problems. Trace model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B (served via vLLM) Label model: gpt-4o-mini Source traces: jrosseruk/DeepSeek-R1-Distill-Llama-8B-MATH-traces-balanced Total sentences: 435,525 Backtrack sentences: 39,998 (9.2%) Traces: 4,413 (2,496 correct, 1,917 incorrect) Each sentence in a chain-of-thought trace is… See the full description on the dataset page: https://huggingface.co/datasets/jrosseruk/DeepSeek-R1-Distill-Llama-8B-MATH-labeled-sentences.tabulartext-generation100K<n<1M0 likes21 downloads8mo agoHugging Face04jrosseruk /DeepSeek-R1-Distill-Llama-8B-MATH-traces-balanced DeepSeek-R1-Distill-Llama-8B MATH Reasoning Traces (Balanced) 4,492 reasoning traces from DeepSeek-R1-Distill-Llama-8B on MATH problems, balanced for correct/incorrect. Model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B (served via vLLM) Source problems: xDAN2099/lighteval-MATH (train split) Sampling: Subsampled from the full 10k trace set — 2,500 correct + 1,992 incorrect (all available incorrect traces) Generation params: temperature=0.6, top_p=0.95, max_tokens=15000 Accuracy: 55.7%… See the full description on the dataset page: https://huggingface.co/datasets/jrosseruk/DeepSeek-R1-Distill-Llama-8B-MATH-traces-balanced.tabulartext-generation1K<n<10K0 likes16 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.