CoolFace
Datasetpublic

ComparEdge/llm-api-benchmark-matrix-2026

LLM Benchmark & Feature Matrix 2026 Which LLM is best at what? This dataset maps capabilities, performance, and limits of 22 major models. Unlike pricing datasets, this focuses on what models can do — not just what they cost. Files File Description llm-benchmarks-2026.csv MMLU, HumanEval, MATH, Arena ELO, coding/reasoning/multilingual rankings, tier (S+ to B) llm-features-2026.csv 15 binary capabilities: vision, function calling, JSON mode… See the full description on the dataset page: https://huggingface.co/datasets/ComparEdge/llm-api-benchmark-matrix-2026.

sourceHugging Facecc-by-4.0updated 5mo agoView on Hugging Face
1likes55downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
ComparEdge/llm-api-benchmark-matrix-2026 · CoolFace