CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ceval /ceval-examC-Eval is a comprehensive Chinese evaluation suite for foundation models. It consists of 13948 multi-choice questions spanning 52 diverse disciplines and four difficulty levels. Please visit our website and GitHub or check our paper for more details. Each subject consists of three splits: dev, val, and test. The dev set per subject consists of five exemplars with explanations for few-shot evaluation. The val set is intended to be used for hyperparameter tuning. And the test set is for model… See the full description on the dataset page: https://huggingface.co/datasets/ceval/ceval-exam.texttext-classification10K<n<100K315 likes147k downloads1y agoHugging Face02mhardalov /exams Dataset Card for [Dataset Name] Dataset Summary EXAMS is a benchmark dataset for multilingual and cross-lingual question answering from high school examinations. It consists of more than 24,000 high-quality high school exam questions in 16 languages, covering 8 language families and 24 school subjects from Natural Sciences and Social Sciences, among others. Supported Tasks and Leaderboards [More Information Needed] Languages The languages in the… See the full description on the dataset page: https://huggingface.co/datasets/mhardalov/exams.textquestion-answering100K<n<1M40 likes24k downloads3y agoHugging Face03cryptom /ceval-examC-Eval is a comprehensive Chinese evaluation suite for foundation models. It consists of 13948 multi-choice questions spanning 52 diverse disciplines and four difficulty levels.texttext-classification10K<n<100K2 likes19k downloads3y agoHugging Face04NTU-yiwen /code-world-model-inference-examples-40 Inference examples This directory contains 40 numbered, independent inference examples. Every example uses only its public number; source case names and internal paths are intentionally omitted. Each numbered directory contains: first_frame.png: exact 1536x864 generated RGB first frame used by inference. prompts/*.txt: the exact rolling long-inference prompts used for the result. condition/*.npz: ordered lossless condition chunks. metadata.json: frame count, FPS, prompt windows… See the full description on the dataset page: https://huggingface.co/datasets/NTU-yiwen/code-world-model-inference-examples-40.text1K<n<10K0 likes4.8k downloads29d agoHugging Face05gchindemi /lisbet-examplestext10K<n<100K0 likes3.8k downloads2y agoHugging Face06chiuratto-AIgourakis /sounio-code-examples Sounio Curated Code Examples Curated compile-clean .sio examples for training and evaluating code models on Sounio, a self-hosted systems and scientific programming language for epistemic computing, uncertainty propagation, and algebraic effects. This directory is the Cx-1 expansion lane for chiuratto-AIgourakis/sounio-code-examples. Current batch Examples: 5,000 Metadata files: 5,000 Compiler gate: bin/souc check pass rate 5,000/5,000 Utility layer: 5,000… See the full description on the dataset page: https://huggingface.co/datasets/chiuratto-AIgourakis/sounio-code-examples.texttext-generation1K<n<10K0 likes2.7k downloads4mo agoHugging Face07JacobLinCool /taiwan-examsMachine-gradable exam benchmarks produced by any-to-bench. Each subset is one exam: the viewer table shows one row per answerable question (figures embedded); the raw, byte-faithful bundle lives under <subset>/bundle/ — exam.json (structured paper), answer_schema.json (strict JSON Schema an answer sheet must satisfy), grading.json (deterministic rules + judge rubrics), manifest.json (provenance), and assets/ (figures). Usage Benchmark any model against an exam: a2b download… See the full description on the dataset page: https://huggingface.co/datasets/JacobLinCool/taiwan-exams.imagequestion-answering1K<n<10K0 likes2.7k downloads1mo agoHugging Face08datasets-examples /doc-formats-jsonl-1 [doc] formats - jsonl - 1 This dataset contains one jsonl file at the root. textn<1K0 likes2.6k downloads3y agoHugging Face09iamroot /chat_formatted_examplestextn<1K0 likes2.4k downloads2y agoHugging Face10matichon /thai-onet-m6-exam Thai O-Net Exams Dataset Overview The Thai O-Net Exams dataset is a comprehensive collection of exam questions and answers from the Thai Ordinary National Educational Test (O-Net). This dataset covers various subjects for Grade 12 (M6) level, designed to assist in educational research and development of question-answering systems. Dataset Source Thai National Institute of Educational Testing Service (NIETS) Maintainer Dr. Kobkrit Viriyayudhakorn… See the full description on the dataset page: https://huggingface.co/datasets/matichon/thai-onet-m6-exam.textquestion-answering1K<n<10K0 likes2.2k downloads5mo agoHugging Face11typhoon-ai /thai_exam Dataset Card for Thai_Exam ThaiExam is a Thai knowledge benchmarking dataset, consisting of multiple-choice questions from examinations in Thailand. The dataset was originally developed for evaluating Typhoon (Thai LLM). This dataset contains 5 splits corresponding to 5 examinations as follows: ONET: The Ordinary National Educational Test (ONET) is an examination for students in Thailand. This dataset is based on the grade-12 ONET exam, comprising 4 subjects and each question has 5… See the full description on the dataset page: https://huggingface.co/datasets/typhoon-ai/thai_exam.tabularquestion-answeringn<1K19 likes2.2k downloads2y agoHugging Face12datasets-examples /doc-formats-parquet-1textn<1K0 likes2.1k downloads2y agoHugging Face13datasets-examples /doc-formats-csv-1 [doc] formats - csv - 1 This dataset contains one csv file at the root: data.csv kind,sound dog,woof cat,meow pokemon,pika human,hello The YAML section of the README does not contain anything related to loading the data (only the size category metadata): --- size_categories: - n<1K --- textn<1K0 likes2.1k downloads3y agoHugging Face14agents-last-exam /agents-last-exam Agents Last Exam — Task Card Metadata (v1.1) A metadata-only release (v1.1) of 152 tasks from the Agents Last Exam (ALE) benchmark for evaluating computer-use agents on long-horizon professional work. The Agents Last Exam dataset family ALE is published as three companion HuggingFace datasets: Dataset Contents Access Task Card Metadata One row per task: titles, prompts, taxonomy, input-file descriptors Open Task Input Data The input/ files each task… See the full description on the dataset page: https://huggingface.co/datasets/agents-last-exam/agents-last-exam.textn<1K210 likes2.1k downloads5d agoHugging Face15Medilora /us_medical_license_exam_textbooks_entext100K<n<1M13 likes1.8k downloads3y agoHugging Face16pgmpy /example_datasetsThis is a collection of example datasets from various sources that the datasets module in pgmpy supports. All credits for creating the datasets go to the original authors. The following datasets are included (along with their LICENCE). The licences are included in the respective dataset folders as well. example-causal-datasets: CC0 1.0 Universal. Last synced on 2026-02-05. nslm: CC0 1.0 Universal. Last downloaded on 2026-03-04. Tuebingen-pair-wise-dataset: Last downloaded on 2026-03-02.… See the full description on the dataset page: https://huggingface.co/datasets/pgmpy/example_datasets.text1K<n<10K0 likes1.7k downloads9d agoHugging Face17openthaigpt /thai-onet-m6-exam Thai O-Net Exams Dataset Overview The Thai O-Net Exams dataset is a comprehensive collection of exam questions and answers from the Thai Ordinary National Educational Test (O-Net). This dataset covers various subjects for Grade 12 (M6) level, designed to assist in educational research and development of question-answering systems. Dataset Source Thai National Institute of Educational Testing Service (NIETS) Maintainer Dr. Kobkrit… See the full description on the dataset page: https://huggingface.co/datasets/openthaigpt/thai-onet-m6-exam.textquestion-answering1K<n<10K8 likes1.6k downloads3y agoHugging Face18R2MED /MedXpertQA-Exam 🔭 Overview R2MED: First Reasoning-Driven Medical Retrieval Benchmark R2MED is a high-quality, high-resolution synthetic information retrieval (IR) dataset designed for medical scenarios. It contains 876 queries with three retrieval tasks, five medical scenarios, and twelve body systems. Dataset #Q #D Avg. Pos Q-Len D-Len Biology 103 57359 3.6 115.2 83.6 Bioinformatics77 47473 2.9 273.8 150.5 Medical Sciences 88 34810 2.8 107.1 122.7 MedXpertQA-Exam 97… See the full description on the dataset page: https://huggingface.co/datasets/R2MED/MedXpertQA-Exam.texttext-retrieval10K<n<100K1 likes1.3k downloads1y agoHugging Face19MemGPT /example-sec-filingstext10K<n<100K7 likes1.3k downloads3y agoHugging Face20erhwenkuo /ceval-exam-zhtw Dataset Card for "ceval-exam-zhtw" C-Eval 是一個針對基礎模型的綜合中文評估套件。它由 13,948 道多項選擇題組成,涵蓋 52 個不同的學科和四個難度級別。原始網站和 GitHub 或查看論文以了解更多詳細資訊。 C-Eval 主要的數據都是使用簡體中文來撰寫并且用來評測簡體中文的 LLM 的效能來設計的,本數據集使用 OpenCC 來進行簡繁的中文轉換,主要目的方便繁中 LLM 的開發與驗測。 下載 使用 Hugging Face datasets 直接載入資料集: from datasets import load_dataset dataset=load_dataset(r"erhwenkuo/ceval-exam-zhtw",name="computer_network") print(dataset['val'][0]) # {'id': 0, 'question':… See the full description on the dataset page: https://huggingface.co/datasets/erhwenkuo/ceval-exam-zhtw.text10K<n<100K0 likes1.2k downloads3y agoHugging Face21OALL /Arabic_EXAMStextn<1K3 likes1.1k downloads3y agoHugging Face22MemGPT /example_short_storiestext1K<n<10K4 likes1k downloads3y agoHugging Face23shashankskagnihotri /humanitys-second-last-exam Humanity's Second Last Exam Benchmark design, curation and release maintenance: Shashank Agnihotri. Original questions retain their recorded authorship and source attribution. This owner-reviewed retained release contains 365 target questions, 730 context examples, and 365 ordered target/A/B links: 1,095 question rows. The owner review concluded on 16 September 2026. This is an owner-reviewed release after suspected-AI-content exclusions, not a software-certified guarantee of… See the full description on the dataset page: https://huggingface.co/datasets/shashankskagnihotri/humanitys-second-last-exam.image1K<n<10K2 likes779 downloads6d agoHugging Face24nielsr /docvqa_1200_examplesimage1K<n<10K12 likes769 downloads4y agoHugging Face25mjbommar /opengloss-v1.3-query-examples-flat See also OpenGloss v2.1 (2026-09-07): a deeper release of 109,633 of these headwords — sense-level ids, four reading levels, sense-tagged examples with spans, a judged relation graph, and retrieval supervision — published as a 16-dataset family. v1.3 remains the broader headword list. OpenGloss Query Examples v1.3 (Flattened) Dataset Summary OpenGloss Query Examples is a synthetic dataset of search queries generated for vocabulary terms. Each term has multiple… See the full description on the dataset page: https://huggingface.co/datasets/mjbommar/opengloss-v1.3-query-examples-flat.texttext-generation100K<n<1M0 likes763 downloads18d agoHugging Face26eduagarcia /oab_examstext1K<n<10K7 likes644 downloads3y agoHugging Face27seedboxai /multitask_german_examples_32ktabular100K<n<1M15 likes560 downloads3y agoHugging Face28MBZUAI /EXAMS-V EXAMS-V: ImageCLEF 2025 – Multimodal Reasoning Dimitar Iliyanov Dimitrov, Hee Ming Shan, Zhuohan Xie, Rocktim Jyoti Das , Momina Ahsan, Sarfraz Ahmad, Nikolay Paev, Ali Mekky, Omar El Herraoui, Rania Hossam, Nurdaulet Mukhituly, Akhmed Sakip, Ivan Koychev, Preslav Nakov INTRODUCTION EXAMS-V is a multilingual, multimodal dataset created to evaluate and benchmark the visual reasoning abilities of AI systems, especially Vision-Language Models (VLMs). The dataset contains… See the full description on the dataset page: https://huggingface.co/datasets/MBZUAI/EXAMS-V.image10K<n<100K10 likes547 downloads1y agoHugging Face29stevenhsu123 /chinese_exam_train_datatext1K<n<10K0 likes542 downloads3y agoHugging Face30dry-melon /Chinese-middle-school-English-exam-questions Dataset Card for Chinese Middle School English Exam Questions If this dataset benefits your work or research, a ❤️ would be greatly appreciated to help others discover it. Dataset Summary A structured collection of English exam questions for Chinese middle school students (grades 7 to 9), including multiple choice, cloze tests, and reading comprehension problems. Supported Tasks and Leaderboards Multiple Choice QA Cloze Test (Multuple Choice, Free Response)… See the full description on the dataset page: https://huggingface.co/datasets/dry-melon/Chinese-middle-school-English-exam-questions.text100K<n<1M2 likes527 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.