kvcache
Datasets
All datasets matching “kvcache”kvcache-quantization-logs-qwen7bKVCacheskvcache-bench-results
Mingxin KV-Cache Tiered-Storage Benchmark Results
Measured results for LLM KV Cache tiered-storage acceleration on 8× AMD Instinct MI308X (ROCm 7.2, vLLM 0.20.1+rocm721, LMCache upstream mainline), model Qwen3-Coder-480B-FP8, published by Mingxin Technology.
Device under test: Mingxin FX100 all-flash NVMe-oF array (4-disk RAID0, RoCEv2, 100 GbE) vs local NVMe (PCIe Gen4) vs recompute-only baseline.
Files
kvcache_bench_results.json — all experiments in structured… See the full description on the dataset page: https://huggingface.co/datasets/wangqiyuan2026/kvcache-bench-results.kv_cache_logits_exp2kvcachekv-cache-compression-mbe
Matched-Budget Evaluation (MBE) — KV Cache Compression
A standardized reporting protocol for KV cache compression in LLM inference. MBE is
not a new task benchmark; it is a thin reporting layer that fixes which models, tasks,
and budgets results are reported at, so that numbers from different papers become
comparable.
Manifest (mbe_manifest.json): the frozen evaluation specification — model suite,
task suite (consuming existing benchmarks: LongBench, RULER, SCBench, GSM8K… See the full description on the dataset page: https://huggingface.co/datasets/Rohithreddybc/kv-cache-compression-mbe.
