datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
open-models-benchmark-results
⚡ Local LLM Evaluation Leaderboard
Welcome to the official public benchmark leaderboard maintained by @ahmedBargady.This dataset repository hosts benchmark evaluation metrics, accuracy scores, throughput telemetry, and quantization trade-off analyses of open-weights foundation models tested locally on NVIDIA A100 GPUs.
💻 Hardware & System Specifications
All evaluations are executed under standardized local cluster environments:
Specification
Details… See the full description on the dataset page: https://huggingface.co/datasets/ahmedBargady/open-models-benchmark-results.eu-open-weight-models
EU-readiness of open-weight LLMs
Curated by LLM Radar — updated 2026-05-03 — 55 models.
A manually-reviewed dataset assessing open-weight Large Language Models (LLMs)
on their suitability for EU deployment and commercial use. Each model is
evaluated on licence, commercial use, training data, and
origin, with quality / speed / price metrics from
Artificial Analysis where available.
Primary use cases:
Selecting open-weight models for self-hosted EU deployment
Licensing and… See the full description on the dataset page: https://huggingface.co/datasets/llmradar/eu-open-weight-models.
