ComparEdge/llm-api-benchmark-matrix-2026
LLM Benchmark & Feature Matrix 2026 Which LLM is best at what? This dataset maps capabilities, performance, and limits of 22 major models. Unlike pricing datasets, this focuses on what models can do — not just what they cost. Files File Description llm-benchmarks-2026.csv MMLU, HumanEval, MATH, Arena ELO, coding/reasoning/multilingual rankings, tier (S+ to B) llm-features-2026.csv 15 binary capabilities: vision, function calling, JSON mode… See the full description on the dataset page: https://huggingface.co/datasets/ComparEdge/llm-api-benchmark-matrix-2026.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face