CoolFace
Datasetpublic

Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca

deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.inference.chatbot-short (NVIDIA H100 80GB HBM3) Back to leaderboard Headline metrics Metric Value Unit TTFT P50 74.4268 ms TTFT P99 521.1922 ms TPOT P50 21.0371 ms TPOT P99 23.458 ms Total P50 Ms 2661.4077 Total P99 Ms 3136.5972 Req Per S Passing 1.0172 Req Per S All 1.1559 Compliance Rate 0.88 Ok Rate 1 Throughput Tok Per S 134.316 Power Avg W 808.3085 Power Peak W… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
0likes18downloads
Dataset Card

deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.inference.chatbot-short (NVIDIA H100 80GB HBM3)

Back to leaderboard

Headline metrics

MetricValueUnit
TTFT P5074.4268ms
TTFT P99521.1922ms
TPOT P5021.0371ms
TPOT P9923.458ms
Total P50 Ms2661.4077
Total P99 Ms3136.5972
Req Per S Passing1.0172
Req Per S All1.1559
Compliance Rate0.88
Ok Rate1
Throughput Tok Per S134.316
Power Avg W808.3085
Power Peak W854.62
Energy Joules Total17249.0462
Joules per token5.9377J
Slo Hardware Classh100
Slo Template Resolvedttft<200ms, tpot<50ms, total<3000ms

Run configuration

  • —Model: deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct @ unknown00
  • —Engine: vllm v0.21.0
  • —Quantization: fp16
  • —Hardware: NVIDIA H100 80GB HBM3
  • —Driver: 580.126.09
  • —CUDA: 13.0
  • —Run date: 2026-05-18T14:13:19.690272+00:00
  • —Seed: 42

Verification

This result is Sigstore-signed and Rekor-logged. Verify:

bash
pip install inferencebench
bench verify hf://datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca/envelope.json

Rekor entry: log index -1

Methodology

See the suite methodology page.

Citation

bibtex
@misc{inferencebench_019e3b6f22ca,
  title = { deepseek-ai/DeepSeek-Coder-V2-Lite-Instruct on llm.inference.chatbot-short },
  author = { {InferenceBench community} },
  year = { 2026 },
  url = { https://huggingface.co/datasets/Yobitel/deepseek-ai-deepseek-coder-v2-lite-instruct__llm-inference-chatbot-short__019e3b6f22ca },
}

Published via [InferenceBench](https://github.com/yobitelcomm/bench) — vendor-neutral AI benchmarks.