Inferencebench/pass-at-k-benchmark-results
ISO-Bench Pass@k GPU Benchmark Results Agent-generated optimization patches benchmarked on real GPU hardware (NVIDIA H100 80GB). Scope: Lossfunk/ISO-Bench — 54 tasks (39 vLLM + 15 SGLang) Patches from: Inferencebench/pass-at-k-samples Summary vLLM SGLang Total Tasks benchmarked 29/39 10/15 39/54 Successful benchmarks 444 150 594 Agent vLLM SGLang Total Claude Code (Sonnet 4.5) 230 76 306 Codex CLI (GPT-5) 214 74 288… See the full description on the dataset page: https://huggingface.co/datasets/Inferencebench/pass-at-k-benchmark-results.
This repository belongs to Inferencebench on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
