CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01huggingface-projects /drlc-leaderboard-datatabular10K<n<100K2 likes33k downloads2d agoHugging Face02hf-audio /open-asr-leaderboard-resultstabularn<1K0 likes5.2k downloads4h agoHugging Face03TIGER-Lab /mmlu_pro_leaderboard_submissiontextn<1K1 likes3.3k downloads7mo agoHugging Face04pkalkman /drlc-leaderboard-datatabular10K<n<100K0 likes961 downloads2y agoHugging Face05hf-audio /leaderboard_longformtabularn<1K0 likes878 downloads3mo agoHugging Face06IqraEval /leaderboard_datatext1K<n<10K0 likes336 downloads26d agoHugging Face07vectara /hhem_leaderboard_datasetstext1K<n<10K0 likes227 downloads1y agoHugging Face08witcheer /agentic-score-leaderboard 🛠️ Agentic Score Leaderboard — one RTX 5090 How well do local models actually drive a tool-using agent loop? Not single-call function-calling benchmarks — a real loop: native OpenAI tool-calling through llama-server, multi-step deterministic tasks, programmatic verification. Everything runs on a single RTX 5090 32GB. Updated 2026-06-17 · llama.cpp b9562 · --jinja native tool-calling · temp 0. Leaderboard # model params Agentic Score success tool-eff… See the full description on the dataset page: https://huggingface.co/datasets/witcheer/agentic-score-leaderboard.tabularn<1K3 likes150 downloads3mo agoHugging Face09mayank-dubey-ai /l4-gpu-llm-benchmark-leaderboard 🚀 Local LLM Serving & Quality Benchmark Leaderboard (NVIDIA L4 24GB) An exhaustive, reproducible benchmark study measuring real-world serving performance (TTFT, TPOT, throughput, peak VRAM, energy consumption, and cost) alongside rigorous task quality gates (HumanEval+, MMLU-Pro, BFCL v4 tool calling, and RULER needle retrieval) for open-weight LLMs on a single NVIDIA L4 24GB GPU. 📊 Executive Summary & Key Takeaways ⚡ Best Throughput & Coding Workhorse:… See the full description on the dataset page: https://huggingface.co/datasets/mayank-dubey-ai/l4-gpu-llm-benchmark-leaderboard.tabulartext-generationn<1K0 likes145 downloads1mo agoHugging Face10oruk /oruk-bench-leaderboard oruk-bench leaderboard Results for 64 speech emotion recognition systems measured on one held-out multilingual evaluation with a single scoring implementation: open checkpoints, closed APIs, audio LLMs, and text-only baselines, all on the same protocol. This dataset is the results table, not the audio. The evaluation clips are assembled from several emotional-speech corpora whose licences differ, so they are not redistributable; the benchmark card documents provenance and how… See the full description on the dataset page: https://huggingface.co/datasets/oruk/oruk-bench-leaderboard.tabularaudio-classificationn<1K1 likes134 downloads14d agoHugging Face11getomni-ai /ocr-leaderboard OmniAI OCR Leaderboard A comprehensive leaderboard comparing OCR and data extraction performance across traditional OCR providers and multimodal LLMs, such as gpt-4o and gemini-2.0. The dataset includes full results from testing 9 providers on 1,000 pages each. Benchmark Results (Feb 2025) | Source Code image1K<n<10K8 likes90 downloads2y agoHugging Face12hivex-research /hivex-leaderboard-datatabularn<1K0 likes66 downloads2y agoHugging Face13Lyte /tokenizer-leaderboard Dataset Card for Dataset Name Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): en License: mit Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed] Uses Direct Use [More… See the full description on the dataset page: https://huggingface.co/datasets/Lyte/tokenizer-leaderboard.tabularn<1K0 likes58 downloads4mo agoHugging Face14mikcnt /cwm-workout-leaderboard-datatextn<1K0 likes48 downloads24d agoHugging Face15huph22 /drlc-leaderboard-datatabular10K<n<100K0 likes43 downloads2y agoHugging Face16vectara /leaderboard_resultstext100K<n<1M5 likes38 downloads1y agoHugging Face17Deddy /leaderboard-datasettabular1K<n<10K0 likes35 downloads1y agoHugging Face18razsarusi /open-llm-leaderboard-eda 🤖 Open LLM Leaderboard – Exploratory Data Analysis Overview This project presents an end-to-end Exploratory Data Analysis (EDA) of the Open LLM Leaderboard dataset from HuggingFace. The goal is to understand what factors predict the overall benchmark performance of open-source Large Language Models (LLMs). The analysis is based on a dataset containing 4,575 LLM evaluation records, including model size, training type, architecture, and scores across 6 standardized… See the full description on the dataset page: https://huggingface.co/datasets/razsarusi/open-llm-leaderboard-eda.tabular1K<n<10K0 likes32 downloads6mo agoHugging Face19wenhu /science_leaderboard_submissionThis dataset contains the results used for Science Leaderboard tabularquestion-answeringn<1K0 likes31 downloads2y agoHugging Face20aliarda /LLMs-Turkish-TEOG-Leaderboard TEOG Scores Leaderboard Welcome to the TEOG Scores Leaderboard! This repository contains the results of evaluating various large language models (LLMs) on the TEOG (Temel Eğitimden Ortaöğretime Geçiş) exam dataset. The TEOG exam is a standardized test in Turkey used for high school admissions, and this dataset provides a benchmark for assessing the performance of LLMs in Turkish educational tasks. Please remember that full score for TEOG is 500 points. More Models Are… See the full description on the dataset page: https://huggingface.co/datasets/aliarda/LLMs-Turkish-TEOG-Leaderboard.textn<1K2 likes23 downloads2y agoHugging Face21PeanutUp /membench_leaderboard_submissiontabularn<1K0 likes23 downloads2mo agoHugging Face22lpmeyer /LLM-KG-Bench-LeaderboardLeaderboard for RDF Knowledge Graph(KG) related capabilities of Large Language Models(LLMs) as generated with the LLM-KG-Bench framework. Results for more than 20 RDF related tasks are collected for more than 40 LLMs. The leaderboard contains summarized results: board_combined_scores_short.csv: the most concise summary, listing for each LLM combined scores in the RDF (R) and SPARQL (S) handling categories, estimating read(R) and write(W) capabilities for syntax(syn) and semantic(sem). If a… See the full description on the dataset page: https://huggingface.co/datasets/lpmeyer/LLM-KG-Bench-Leaderboard.textn<1K0 likes23 downloads3mo agoHugging Face23gberseth /rl-leaderboard-requeststextn<1K0 likes21 downloads5mo agoHugging Face24AIDX-ktds /ko_leaderboard한국어 리더보드 학습에서 사용된 데이터 중 약 8,000건에 대해서 공개합니다. 데이터 생성 시 도움이 되기를 바랍니다. 감사합니다. texttext-generation1K<n<10K4 likes18 downloads2y agoHugging Face25med-llm-leaderboard /shadrtextn<1K0 likes17 downloads3y agoHugging Face26reach-vb /open-asr-leaderboard-evals-alltabularn<1K0 likes16 downloads3y agoHugging Face27autorl-org /arlbench-leaderboard-results ARLBench Leaderboard Data This is the data repository for the ARLBench leaderboard. If you want to add your own runs to the leaderboard, please make a PR here with the following content: The data file(s) with your data. There's a template available you can fill in. If you add data for all algorithms, you can use a single file. If you're only adding for a subset, add separate files per algorithm. The extended data split. This means extending the ReadMe.md config. If you added a file… See the full description on the dataset page: https://huggingface.co/datasets/autorl-org/arlbench-leaderboard-results.tabularn<1K0 likes13 downloads11mo agoHugging Face28Meo-Advisors /ai-adoption-leaderboards AI Adoption Leaderboards (US Geography & Industry) 6,822 pre-computed leaderboards ranking AI adoption across all 50 US states, 300+ metros, 1,200+ cities, and 100+ NAICS industries — with company counts, average scores, and top companies per segment. Rows: 6,822 Source: Derived from the AI Adoption Index (US Companies) Methodology + interactive explorer: https://meoadvisors.com/ai-opportunities/leaderboard/ Part of: the open AI Workforce Data collection by Meo Advisors… See the full description on the dataset page: https://huggingface.co/datasets/Meo-Advisors/ai-adoption-leaderboards.tabular1K<n<10K0 likes12 downloads4mo agoHugging Face29reach-vb /open-asr-leaderboard-evals-ex-cvtabularn<1K0 likes10 downloads3y agoHugging Face30airlsyn /leaderboard_datasettext100K<n<1M0 likes9 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.