Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.inference.chatbot-short (NVIDIA H100 80GB HBM3) Back to leaderboard Headline metrics Metric Value Unit TTFT P50 74.4268 ms TTFT P99 521.1922 ms TPOT P50 21.0371 ms TPOT P99 23.458 ms Total P50 Ms 2661.4077 Total P99 Ms 3136.5972 Req Per S Passing 1.0172 Req Per S All 1.1559 Compliance Rate 0.88 Ok Rate 1 Throughput Tok Per S 134.316 Power Avg W 808.3085 Power Peak W… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca.
deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.inference.chatbot-short (NVIDIA H100 80GB HBM3)
Headline metrics
Run configuration
- Model: deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct @ unknown00
- Engine: vllm v0.21.0
- Quantization: fp16
- Hardware: NVIDIA H100 80GB HBM3
- Driver: 580.126.09
- CUDA: 13.0
- Run date: 2026-05-18T14:13:19.690272+00:00
- Seed: 42
Verification
This result is Sigstore-signed and Rekor-logged. Verify:
pip install inferencebench
bench verify hf://datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca/envelope.jsonMethodology
See the suite methodology page.
Citation
@misc{inferencebench_019e3b6f22ca,
title = { deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.inference.chatbot-short },
author = { {InferenceBench community} },
year = { 2026 },
url = { https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca },
}Published via [InferenceBench](https://github.com/yobitelcomm/bench) — vendor-neutral AI benchmarks.
