CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01huggingface-projects /drlc-leaderboard-datatabular10K<n<100K2 likes33k downloads1d agoHugging Face02hf-audio /open-asr-leaderboard-resultstabularn<1K0 likes5.2k downloads2d agoHugging Face03pkalkman /drlc-leaderboard-datatabular10K<n<100K0 likes961 downloads2y agoHugging Face04hf-audio /leaderboard_longformtabularn<1K0 likes878 downloads3mo agoHugging Face05EleutherAI /bergson-wikitext-gpt2-leaderboard-bank bergson leaderboard: retrain banks, scores and LDS/QLD results (WikiText GPT-2) Everything behind the numbers on the bergson leaderboard, for the model at EleutherAI/bergson-wikitext-gpt2-leaderboard. path what it is bank/ the LDS ground truth: 100 random leave-1%-out subsets of the 4,608 training chunks (subsets.json) and each subset's measured loss change on the 50 test queries (validation.csv) random/retrained/{base,subset_0..99} the retrained models themselves… See the full description on the dataset page: https://huggingface.co/datasets/EleutherAI/bergson-wikitext-gpt2-leaderboard-bank.tabular10K<n<100K0 likes179 downloads9d agoHugging Face06witcheer /agentic-score-leaderboard 🛠️ Agentic Score Leaderboard — one RTX 5090 How well do local models actually drive a tool-using agent loop? Not single-call function-calling benchmarks — a real loop: native OpenAI tool-calling through llama-server, multi-step deterministic tasks, programmatic verification. Everything runs on a single RTX 5090 32GB. Updated 2026-06-17 · llama.cpp b9562 · --jinja native tool-calling · temp 0. Leaderboard # model params Agentic Score success tool-eff… See the full description on the dataset page: https://huggingface.co/datasets/witcheer/agentic-score-leaderboard.tabularn<1K3 likes150 downloads3mo agoHugging Face07mayank-dubey-ai /l4-gpu-llm-benchmark-leaderboard 🚀 Local LLM Serving & Quality Benchmark Leaderboard (NVIDIA L4 24GB) An exhaustive, reproducible benchmark study measuring real-world serving performance (TTFT, TPOT, throughput, peak VRAM, energy consumption, and cost) alongside rigorous task quality gates (HumanEval+, MMLU-Pro, BFCL v4 tool calling, and RULER needle retrieval) for open-weight LLMs on a single NVIDIA L4 24GB GPU. 📊 Executive Summary & Key Takeaways ⚡ Best Throughput & Coding Workhorse:… See the full description on the dataset page: https://huggingface.co/datasets/mayank-dubey-ai/l4-gpu-llm-benchmark-leaderboard.tabulartext-generationn<1K0 likes145 downloads1mo agoHugging Face08oruk /oruk-bench-leaderboard oruk-bench leaderboard Results for 64 speech emotion recognition systems measured on one held-out multilingual evaluation with a single scoring implementation: open checkpoints, closed APIs, audio LLMs, and text-only baselines, all on the same protocol. This dataset is the results table, not the audio. The evaluation clips are assembled from several emotional-speech corpora whose licences differ, so they are not redistributable; the benchmark card documents provenance and how… See the full description on the dataset page: https://huggingface.co/datasets/oruk/oruk-bench-leaderboard.tabularaudio-classificationn<1K1 likes134 downloads14d agoHugging Face09getomni-ai /ocr-leaderboard OmniAI OCR Leaderboard A comprehensive leaderboard comparing OCR and data extraction performance across traditional OCR providers and multimodal LLMs, such as gpt-4o and gemini-2.0. The dataset includes full results from testing 9 providers on 1,000 pages each. Benchmark Results (Feb 2025) | Source Code image1K<n<10K8 likes90 downloads2y agoHugging Face10hivex-research /hivex-leaderboard-datatabularn<1K0 likes66 downloads2y agoHugging Face11Lyte /tokenizer-leaderboard Dataset Card for Dataset Name Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): en License: mit Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed] Uses Direct Use [More… See the full description on the dataset page: https://huggingface.co/datasets/Lyte/tokenizer-leaderboard.tabularn<1K0 likes58 downloads4mo agoHugging Face12huph22 /drlc-leaderboard-datatabular10K<n<100K0 likes43 downloads2y agoHugging Face13Deddy /leaderboard-datasettabular1K<n<10K0 likes35 downloads1y agoHugging Face14razsarusi /open-llm-leaderboard-eda 🤖 Open LLM Leaderboard – Exploratory Data Analysis Overview This project presents an end-to-end Exploratory Data Analysis (EDA) of the Open LLM Leaderboard dataset from HuggingFace. The goal is to understand what factors predict the overall benchmark performance of open-source Large Language Models (LLMs). The analysis is based on a dataset containing 4,575 LLM evaluation records, including model size, training type, architecture, and scores across 6 standardized… See the full description on the dataset page: https://huggingface.co/datasets/razsarusi/open-llm-leaderboard-eda.tabular1K<n<10K0 likes32 downloads6mo agoHugging Face15wenhu /science_leaderboard_submissionThis dataset contains the results used for Science Leaderboard tabularquestion-answeringn<1K0 likes31 downloads2y agoHugging Face16PeanutUp /membench_leaderboard_submissiontabularn<1K0 likes23 downloads2mo agoHugging Face17reach-vb /open-asr-leaderboard-evals-alltabularn<1K0 likes16 downloads3y agoHugging Face18autorl-org /arlbench-leaderboard-results ARLBench Leaderboard Data This is the data repository for the ARLBench leaderboard. If you want to add your own runs to the leaderboard, please make a PR here with the following content: The data file(s) with your data. There's a template available you can fill in. If you add data for all algorithms, you can use a single file. If you're only adding for a subset, add separate files per algorithm. The extended data split. This means extending the ReadMe.md config. If you added a file… See the full description on the dataset page: https://huggingface.co/datasets/autorl-org/arlbench-leaderboard-results.tabularn<1K0 likes13 downloads11mo agoHugging Face19Meo-Advisors /ai-adoption-leaderboards AI Adoption Leaderboards (US Geography & Industry) 6,822 pre-computed leaderboards ranking AI adoption across all 50 US states, 300+ metros, 1,200+ cities, and 100+ NAICS industries — with company counts, average scores, and top companies per segment. Rows: 6,822 Source: Derived from the AI Adoption Index (US Companies) Methodology + interactive explorer: https://meoadvisors.com/ai-opportunities/leaderboard/ Part of: the open AI Workforce Data collection by Meo Advisors… See the full description on the dataset page: https://huggingface.co/datasets/Meo-Advisors/ai-adoption-leaderboards.tabular1K<n<10K0 likes12 downloads4mo agoHugging Face20reach-vb /open-asr-leaderboard-evals-ex-cvtabularn<1K0 likes10 downloads3y agoHugging Face21nnagesh101 /memisislabs-leaderboardtabularn<1K0 likes9 downloads2mo agoHugging Face22Steveeeeeeen /leaderboard_evalstabularn<1K0 likes7 downloads1y agoHugging Face23cwchen-cm /leaderboard-testtabularn<1K0 likes6 downloads2y agoHugging Face24aborowska /DSPT-Leaderboard-Resultstabularn<1K0 likes6 downloads3mo agoHugging Face25TIGER-Lab /LongICL_leaderboard_submissiontabularn<1K0 likes5 downloads2y agoHugging Face26Abhyudaya101 /Irish_LLM_Leaderboardtabularn<1K0 likes5 downloads9mo agoHugging Face27SpX-DAC /combined_leaderboard_with_pdf_scorestabularn<1K0 likes5 downloads5mo agoHugging Face28Romihi50 /minicar-leaderboard 🏆 MiniCar Competition Leaderboard 自動運転ミニカーコンペティションのリーダーボードデータ データ構造 カラム 説明 submission_id 提出ID team_name チーム名 model_name モデル名 lap_time ラップタイム(秒) success_rate 成功率(%) track_name トラック名 video_url 動画URL model_repo モデルリポジトリ notes 備考 submitted_at 提出日時 verified 検証済みフラグ 使い方 import pandas as pd from huggingface_hub import hf_hub_download path = hf_hub_download( repo_id="Romihi50/minicar-leaderboard"… See the full description on the dataset page: https://huggingface.co/datasets/Romihi50/minicar-leaderboard.tabularothern<1K0 likes4 downloads9mo agoHugging Face29lisdfdf /open-asr-leaderboard-evals-alltabularn<1K0 likes4 downloads4mo agoHugging Face30gberseth /rl-leaderboard-resultstabularn<1K0 likes3 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.