CoolFace
Datasetpublic

Yobitel/meta-llama-llama-3-1-70b-instruct__code-generation-humaneval-mini__019e3b935675

meta-llama/Llama-3.1-70B-Instruct on code.generation.humaneval-mini (NVIDIA H100 80GB HBM3) Back to leaderboard Headline metrics Metric Value Unit N Samples 5 N Ok 5 Ok Rate 1 Pass At 1 0.8 Pass At 1 P05 0.2 Pass At 1 P50 1 Pass At 1 P95 1 Timeout Rate 0 TTFT P50 28.0842 ms Total P50 Ms 4338.3245 Tokens Out Total 1653 Run configuration Model: meta-llama/Llama-3.1-70B-Instruct @ unknown00 Engine: vllm… See the full description on the dataset page: https://huggingface.co/datasets/Yobitel/meta-llama-llama-3-1-70b-instruct__code-generation-humaneval-mini__019e3b935675.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
0likes17downloads
Dataset Card

meta-llama/Llama-3.1-70B-Instruct on code.generation.humaneval-mini (NVIDIA H100 80GB HBM3)

Back to leaderboard

Headline metrics

MetricValueUnit
N Samples5
N Ok5
Ok Rate1
Pass At 10.8
Pass At 1 P050.2
Pass At 1 P501
Pass At 1 P951
Timeout Rate0
TTFT P5028.0842ms
Total P50 Ms4338.3245
Tokens Out Total1653

Run configuration

  • —Model: meta-llama/Llama-3.1-70B-Instruct @ unknown00
  • —Engine: vllm vunknown
  • —Quantization: fp16
  • —Hardware: NVIDIA H100 80GB HBM3
  • —Driver: 580.126.09
  • —CUDA: 13.0
  • —Run date: 2026-05-18T14:52:52.213409+00:00
  • —Seed: 0

Verification

This result is Sigstore-signed and Rekor-logged. Verify:

bash
pip install inferencebench
bench verify hf://datasets/Yobitel/meta-llama-llama-3-1-70b-instruct__code-generation-humaneval-mini__019e3b935675/envelope.json

Rekor entry: log index -1

Methodology

See the suite methodology page.

Citation

bibtex
@misc{inferencebench_019e3b935675,
  title = { meta-llama/Llama-3.1-70B-Instruct on code.generation.humaneval-mini },
  author = { {InferenceBench community} },
  year = { 2026 },
  url = { https://huggingface.co/datasets/Yobitel/meta-llama-llama-3-1-70b-instruct__code-generation-humaneval-mini__019e3b935675 },
}

Published via [InferenceBench](https://github.com/yobitelcomm/bench) — vendor-neutral AI benchmarks.