CoolFace
14 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01amalia-llm /pt_exams PHEB - Portuguese High School Exams MCQ MCQ set of PHEB a collection of Portuguese exam questions for evaluating language models on academic knowledge on the Portuguese curriculum. For more details, see the PHEB paper. This dataset is provided as part of the AMALIA project and is included in AMALIA-Bench, a comprehensive benchmark suite for evaluating large language models on European Portuguese. Citation If you use this dataset or AMALIA in your work… See the full description on the dataset page: https://huggingface.co/datasets/amalia-llm/pt_exams.tabularquestion-answering1K<n<10K0 likes160 downloads3mo agoHugging Face02bbidpa /flutter-full-examples-v1 Flutter Codegen: Full Examples Synthetic dataset of complete Flutter/Dart widgets, each paired with the goal that describes them and (optionally) starting code. Unlike flutter-codegen-diff-steps, there's no step history or diff structure here -- each row is a single, standalone goal -> complete file example. This is the whole-code counterpart to flutter-diff-steps-v1, intended for training/evaluating a baseline that generates the entire file in one shot, to compare against the… See the full description on the dataset page: https://huggingface.co/datasets/bbidpa/flutter-full-examples-v1.tabulartext-generation10K<n<100K0 likes135 downloads16d agoHugging Face03mpvasilis /thema-panhellenic-exams Thema Thema is an open benchmark based on the Greek Panhellenic university entrance examinations (Πανελλαδικές Εξετάσεις ΓΕΛ). Models answer authentic exam questions in Greek. Responses are graded on the national 0 to 20 scale and can be converted to the admission points used by Greek university departments. Website · Code · Leaderboard · Method Dataset summary Current release Exam years 2023 to 2026 Published subjects 10 of 10 Complete exam… See the full description on the dataset page: https://huggingface.co/datasets/mpvasilis/thema-panhellenic-exams.tabularquestion-answeringn<1K0 likes130 downloads2mo agoHugging Face04neurocheckout-ai /synthetic-abandoned-cart-email-examples Synthetic Abandoned Cart Email Examples An entirely synthetic, bilingual collection of abandoned-cart email drafts with transparent checklist annotations. It is intended for education, prototyping, and evaluation, and contains no real recipients, customer messages, orders, merchant data, or campaign results. Dataset Description The dataset mirrors the five visible checks in NeuroCheckout's public Abandoned Cart Email Checker: message clarity; primary call to… See the full description on the dataset page: https://huggingface.co/datasets/neurocheckout-ai/synthetic-abandoned-cart-email-examples.tabulartext-classificationn<1K0 likes60 downloads24d agoHugging Face05mjbommar /opengloss-v1.3-contrastive-examples See also OpenGloss v2.1 (2026-09-07): a deeper release of 109,633 of these headwords — sense-level ids, four reading levels, sense-tagged examples with spans, a judged relation graph, and retrieval supervision — published as a 16-dataset family. v1.3 remains the broader headword list. OpenGloss Contrastive Examples v1.3 Dataset Summary OpenGloss Contrastive Examples is a synthetic dataset of graduated semantic variations designed for contrastive learning and… See the full description on the dataset page: https://huggingface.co/datasets/mjbommar/opengloss-v1.3-contrastive-examples.tabulartext-generation100K<n<1M0 likes51 downloads18d agoHugging Face06chongpangnasilemak /icd10cm-exam-openended ICD-10-CM Exam Open-Ended 70 open-ended ICD-10-CM coding items with published answer keys. The model writes the answer; there are no options to choose from. The questions are third-party coursework of unverified provenance, reproduced verbatim. The copyright holder is unknown and no licence was granted. If you are the rights holder and object, the dataset will be removed. No certified medical coder or clinician reviewed any item. The answer keys are the coursework's own, not… See the full description on the dataset page: https://huggingface.co/datasets/chongpangnasilemak/icd10cm-exam-openended.tabulartext-generationn<1K0 likes45 downloads23d agoHugging Face07Not-Humanity-Exam /Imprint-Train-v3tabulartext-classification100K<n<1M0 likes33 downloads7mo agoHugging Face08SPAISS6F1 /spai-ss6-corpus-thai-exam-qa-answers SPAI SS6 Thai Exam QA With Answers Index Index repo for normalized Thai O-NET and exam question-answer records with answer keys. This is a lightweight index dataset repo. It does not duplicate the full corpus. The full Parquet data lives in the canonical repository config below. Canonical Data Canonical repo: SPAISS6F1/spai-ss6-llm-1b-thai-corpus Canonical config: thai_exam_qa_with_answers Rows in canonical config: 8,191 Parquet size in canonical config: 0.01 GB… See the full description on the dataset page: https://huggingface.co/datasets/SPAISS6F1/spai-ss6-corpus-thai-exam-qa-answers.tabulartext-generationn<1K0 likes23 downloads4mo agoHugging Face09airesearch /thai-bar-exam-judging Thai Bar-Exam Judging Corpus Anonymised free-form Thai legal essays from a bar-exam preparation exercise, with three Bar Council-trained examiners scoring every essay and span-anchored inline commentary on roughly two thirds of the answers. Eight LLM examinees took the same exam under the same conditions; their answers were graded blind by the same examiners. Fifteen of the 150 answers were cross-graded by the two non-primary examiners, producing the 3-rater stability subset that… See the full description on the dataset page: https://huggingface.co/datasets/airesearch/thai-bar-exam-judging.tabulartext-classification1K<n<10K0 likes19 downloads4mo agoHugging Face10Not-Humanity-Exam /Imprint-Train-v2tabulartext-classification10K<n<100K0 likes17 downloads7mo agoHugging Face11mjbommar /opengloss-v1.1-contrastive-examples OpenGloss Contrastive Examples v1.1 Dataset Summary OpenGloss Contrastive Examples is a synthetic dataset of graduated semantic variations designed for contrastive learning and semantic similarity training. Each example contains a source sentence and a 5-point semantic gradient showing how meaning shifts from antonym to synonym poles. This dataset is derived from the OpenGloss encyclopedic dictionary, using example sentences and their lexical context to generate… See the full description on the dataset page: https://huggingface.co/datasets/mjbommar/opengloss-v1.1-contrastive-examples.tabulartext-generation10K<n<100K0 likes16 downloads10mo agoHugging Face12Not-Humanity-Exam /Imprint-Train-v1tabulartext-classification1K<n<10K0 likes16 downloads7mo agoHugging Face13mjbommar /opengloss-v1.2-contrastive-examples OpenGloss Contrastive Examples v1.2 Dataset Summary OpenGloss Contrastive Examples is a synthetic dataset of graduated semantic variations designed for contrastive learning and semantic similarity training. Each example contains a source sentence and a 5-point semantic gradient showing how meaning shifts from antonym to synonym poles. This dataset is derived from the OpenGloss encyclopedic dictionary, using example sentences and their lexical context to generate… See the full description on the dataset page: https://huggingface.co/datasets/mjbommar/opengloss-v1.2-contrastive-examples.tabulartext-generation10K<n<100K0 likes14 downloads6mo agoHugging Face14ReexpressAI /OpenVerification1_aux_adaptation_examples Dataset Card for ReexpressAI/OpenVerification1_aux_adaptation_examples This is additional data as part of ReexpressAI/OpenVerification1. The data fields are slightly different for this data source, so we include this as a separate dataset. This is example output from the Reexpress MCP Server when using the ReexpressAddTrue, ReexpressAddFalse, or ReexpressAddOOD tools. These are the lines that get saved to the adaptation/running_updates.jsonl file in the model directory. Refer to… See the full description on the dataset page: https://huggingface.co/datasets/ReexpressAI/OpenVerification1_aux_adaptation_examples.tabulartext-classificationn<1K0 likes9 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.