datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
eagle3-speculative-decoding-energy-sweep
EAGLE3 Speculative Decoding Energy Sweep
Per-config energy/throughput/latency measurements for EAGLE3 speculative decoding
(speculative_num_steps, speculative_eagle_topk, speculative_num_draft_tokens)
served with sglang, across batch sizes. Collected for an RL project that learns to
pick speculative-decoding parameters to hold GPU energy utilization in a target band.
Model: unsloth/Llama-3.2-1B-Instruct + rescommons/SpecForge-EAGLE3-Llama-3.2-1B-Instruct draft head.
Hardware:… See the full description on the dataset page: https://huggingface.co/datasets/Pradheep1647/eagle3-speculative-decoding-energy-sweep.speculative-decoding-papers
Speculative Decoding Papers — FineSet
A research-paper dataset on Speculative Decoding Papers, assembled, deduplicated, and quality-scored by
FineSet from arXiv and Semantic Scholar.
📸 This is a dated snapshot — generated 2026-06-19.
It is not auto-updated. Research on Speculative Decoding Papers moves fast — new papers land on arXiv every
week. Want this same dataset refreshed daily, on a topic you choose? See the bottom. ↓
Why this dataset
Quality-scored:… See the full description on the dataset page: https://huggingface.co/datasets/fineset-io/speculative-decoding-papers.Qwen3-speculative-pair-report
Qwen3 speculative pair report: 0.6B draft + 8B target, measured acceptance
Research evidence dataset. No model weights. Part of the collection
Xyntetik Research: Runner Compatibility Reports on this account, produced with
Xyntetik Runner.
Dataset summary
Question tested. What the acceptance rate of a Qwen3-0.6B draft against a Qwen3-8B target actually is across draft depths, whether the engine's printed tokens-per-round figure can be tuned on (it cannot), and… See the full description on the dataset page: https://huggingface.co/datasets/Joakimpalm-Zen/Qwen3-speculative-pair-report.p12-p13-speculative-served-runtime-contract
P12/P13 Speculative Served Runtime Contract
This dataset repo is a GetGenetica-owned benchmark contract package for the
P12/P13 speculative decoding, long-context route, and diffusion-block runtime
lane. It intentionally does not contain a measured served-runtime result.
Current state:
Contract ID: p12_p13_same_checkpoint_served_runtime_contract
Benchmark scope: same_checkpoint_same_hardware_same_prompt_cohort
Result state: blocked_missing_same_checkpoint_served_runtime_result… See the full description on the dataset page: https://huggingface.co/datasets/GetGenetica/p12-p13-speculative-served-runtime-contract.
