datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
speculative-reasoning-matheval-4b-sweep
MathEval sweep — speculative reasoning on Qwen3-4B
Per-sample generations and grading for four arms of a MathEval run, measuring what
speculative reasoning costs and saves against a base model that does not speculate.
Code and write-up: yurun-yuan/speculative-reasoning
— see docs/05-rl-4b.md.
The arms
All four answer the same 1,547 MathEval problems under a 20,000-token response budget.
split
model
runtime
base_plain
Qwen/Qwen3-4B
plain — no speculation… See the full description on the dataset page: https://huggingface.co/datasets/yyuan244/speculative-reasoning-matheval-4b-sweep.speculative-decoding-papers
Speculative Decoding Papers — FineSet
A research-paper dataset on Speculative Decoding Papers, assembled, deduplicated, and quality-scored by
FineSet from arXiv and Semantic Scholar.
📸 This is a dated snapshot — generated 2026-06-19.
It is not auto-updated. Research on Speculative Decoding Papers moves fast — new papers land on arXiv every
week. Want this same dataset refreshed daily, on a topic you choose? See the bottom. ↓
Why this dataset
Quality-scored:… See the full description on the dataset page: https://huggingface.co/datasets/fineset-io/speculative-decoding-papers.
