datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
deepseek-r1-systems-kernel-reasoning
🧠 DeepSeek-R1 Low-Level Systems & Kernel Reasoning Suite (2026)
🛒 Commercial Full Suite Available:
The full production suite with 10,000 SFT Hardware Reasoning Traces + 2,500 High-Contrast DPO Alignment Pairs across all 20 domains is available on Gumroad:
👉 Download Full Commercial Dataset on Gumroad (Starter \ / Pro \ / Enterprise )
A Tier-1 Commercial Dataset Suite engineered specifically for fine-tuning DeepSeek-R1, DeepSeek-R1-Distill-Qwen-14B/32B, and frontier… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/deepseek-r1-systems-kernel-reasoning.DeepSeek-R1-Distill-Llama-8B-MATH-traces
DeepSeek-R1-Distill-Llama-8B MATH Reasoning Traces
10,000 reasoning traces from DeepSeek-R1-Distill-Llama-8B on MATH problems.
Model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B (served via vLLM)
Source problems: xDAN2099/lighteval-MATH (train split)
Sampling: 500 problems (100 per difficulty level 1-5) x 20 rollouts
Generation params: temperature=0.6, top_p=0.95, max_tokens=15000
Accuracy: 80.1% (8,008 correct / 1,992 incorrect)
Problem types: Algebra, Counting & Probability… See the full description on the dataset page: https://huggingface.co/datasets/jrosseruk/DeepSeek-R1-Distill-Llama-8B-MATH-traces.DeepSeek-R1-Distill-Llama-8B-MATH-labeled-sentences
DeepSeek-R1-Distill-Llama-8B MATH Labeled Sentences
Sentence-level function-tag labels for reasoning traces from DeepSeek-R1-Distill-Llama-8B on MATH problems.
Trace model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B (served via vLLM)
Label model: gpt-4o-mini
Source traces: jrosseruk/DeepSeek-R1-Distill-Llama-8B-MATH-traces-balanced
Total sentences: 435,525
Backtrack sentences: 39,998 (9.2%)
Traces: 4,413 (2,496 correct, 1,917 incorrect)
Each sentence in a chain-of-thought trace is… See the full description on the dataset page: https://huggingface.co/datasets/jrosseruk/DeepSeek-R1-Distill-Llama-8B-MATH-labeled-sentences.DeepSeek-R1-Distill-Llama-8B-MATH-traces-balanced
DeepSeek-R1-Distill-Llama-8B MATH Reasoning Traces (Balanced)
4,492 reasoning traces from DeepSeek-R1-Distill-Llama-8B on MATH problems, balanced for correct/incorrect.
Model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B (served via vLLM)
Source problems: xDAN2099/lighteval-MATH (train split)
Sampling: Subsampled from the full 10k trace set — 2,500 correct + 1,992 incorrect (all available incorrect traces)
Generation params: temperature=0.6, top_p=0.95, max_tokens=15000
Accuracy: 55.7%… See the full description on the dataset page: https://huggingface.co/datasets/jrosseruk/DeepSeek-R1-Distill-Llama-8B-MATH-traces-balanced.
