CoolFace
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01WPRM /human_annotation_web_rm_version_1{ "total_stats": { "total_action_data": 19016, "total_website": 50, "action_type": { "bid": 852, "coord": 848 }, "level": { "easy": 486, "medium": 856, "hard": 358 }, "viewport_type": { "full": 707, "laptop": 698, "mobile": 295 }, "judge_type": { "string_match": 858, "url_match": 836… See the full description on the dataset page: https://huggingface.co/datasets/WPRM/human_annotation_web_rm_version_1.image10K<n<100K0 likes261 downloads2y agoHugging Face02Vchitect /VBench-I2V_human_annotationtextn<1K0 likes152 downloads6d agoHugging Face03emarro /example_10kbp_human_annotationstabular100K<n<1M0 likes105 downloads1y agoHugging Face04emarro /example_eval_only_10kb_human_annotationstabular10K<n<100K0 likes83 downloads1y agoHugging Face05izi-ano /CounselBench-Adv-human-annotationtext1K<n<10K0 likes81 downloads5mo agoHugging Face06WRBench /wrbench-human-annotations WRBench Human Annotations This dataset contains the human comparison labels used to validate WRBench's automatic evaluation metrics. Version Update: 2026-07-07 We updated the release after rechecking videos that changed during benchmark maintenance. The release now includes: 1,741 clean comparison rows. 4,302 individual human judgments. 585 newly rechecked current-benchmark comparisons, each reviewed by three annotators. Majority-label summaries for the newly… See the full description on the dataset page: https://huggingface.co/datasets/WRBench/wrbench-human-annotations.tabularimage-to-video1K<n<10K0 likes71 downloads2mo agoHugging Face07JQL-AI /JQL-Human-Edu-Annotations 📚 JQL Multilingual Educational Quality Annotations This dataset provides high-quality human annotations for evaluating the educational value of web documents, and serves as a benchmark for training and evaluating multilingual LLM annotators as described in the JQL paper. 📝 Dataset Summary Documents: 511 English texts Annotations: 3 human ratings per document (0–5 scale) Translations: Into 35 European languages using DeepL and GPT-4o Purpose: For training and… See the full description on the dataset page: https://huggingface.co/datasets/JQL-AI/JQL-Human-Edu-Annotations.texttext-classification10K<n<100K5 likes62 downloads1y agoHugging Face08vishnu2308 /aya23-human-annotationstabular10K<n<100K0 likes33 downloads2y agoHugging Face09Experimental-Orange /HumanAgencyBench_Human_Annotations Human annotations and LLM judge comparative Dataset Paper: HumanAgencyBench: Scalable Evaluation of Human Agency Support in AI Assistants Code: https://github.com/BenSturgeon/HumanAgencyBench/ Dataset Description This dataset contains 60,000 evaluated AI assistant responses across 6 dimensions of behaviour relevant to human agency support, with both model-based and human annotations. Each example includes evaluations from 4 different frontier LLM models. We also provide… See the full description on the dataset page: https://huggingface.co/datasets/Experimental-Orange/HumanAgencyBench_Human_Annotations.texttext-generation10K<n<100K0 likes25 downloads1y agoHugging Face10zjhhhh /human-annotation-1.5B_judge_preference_ternarytextn<1K0 likes24 downloads1y agoHugging Face11gigant /robust_long_abstractive_human_annotationOriginal repository How Far are We from Robust Long Abstractive Summarization? (EMNLP 2022) [Paper] Huan Yee Koh*, Jiaxin Ju*, He Zhang, Ming Liu, Shirui Pan (* denotes equal contribution) Human Annotation of Model-Generated Summaries Data Field Definition dataset Whether the model-generated summary is from arXiv or GovReport dataset. dataset_id ID_ + document ID of the dataset. To match the IDs with original datasets, please remove the "ID_"… See the full description on the dataset page: https://huggingface.co/datasets/gigant/robust_long_abstractive_human_annotation.tabularn<1K0 likes17 downloads2y agoHugging Face12connections-dev /human_annotation_creativitytextn<1K0 likes16 downloads6mo agoHugging Face13thoughtworks /psychometric_human_annotationstextn<1K0 likes14 downloads5mo agoHugging Face14Synthyra /GO_ANNOTATIONS_HUMANhttps://www.uniprot.org/uniprotkb?query=%28go_manual%3A*%29+AND+%28taxonomy_id%3A9606%29 on 1/22/2026 text10K<n<100K0 likes11 downloads8mo agoHugging Face15zjhhhh /human-annotation-1.5Btextn<1K0 likes6 downloads1y agoHugging Face16zjhhhh /human-annotationtextn<1K0 likes5 downloads1y agoHugging Face17ariefansclub /han-human-task-clarity-annotations-v1 Human Task Clarity Annotations This dataset annotates household instructions based on how clear and unambiguous they are for humanoid robotic systems. Annotation Goal To support research on instruction clarity and misinterpretation reduction. Use Cases Instruction parsing Language grounding Assistive robotics research Part of Humanoid Network (HAN) License MIT textn<1K0 likes5 downloads8mo agoHugging Face18zjhhhh /human-annotation-1.5B_judge_preference_ternary_2textn<1K0 likes4 downloads1y agoHugging Face19zjhhhh /human-annotation-1.5B_judge_preference_5score_2textn<1K0 likes4 downloads1y agoHugging Face20MisDrifter /new-human-annotationtextn<1K0 likes4 downloads1y agoHugging Face21rntc /human-annotationstextn<1K0 likes3 downloads2y agoHugging Face22ferocious-aardvark /HumanAgencyBench_Human_Annotations Human annotations and LLM judge comparative Dataset Dataset Description This dataset contains 60,000 evaluated AI assistant responses across 6 dimensions of behaviour relevant to human agency support, with both model-based and human annotations. Each example includes evaluations from 4 different frontier LLM models. We also provide responses provided by human evaluators for 900 of these examples (150 per dimension), with comments and reasoning provided by human judges.… See the full description on the dataset page: https://huggingface.co/datasets/ferocious-aardvark/HumanAgencyBench_Human_Annotations.text10K<n<100K0 likes2 downloads1y agoHugging Face23WPRM /human_annotation_web_rm_tmpgated Dataset Card for "human_annotation_web_rm_tmp" More Information needed image1K<n<10K0 likes1 downloads2y agoHugging Face24anon34957 /HumanAgencyEval_Human_Annotations Human annotations and LLM judge comparative Dataset Dataset Description This dataset contains 60,000 evaluated AI assistant responses across 6 dimensions of behaviour relevant to human agency support, with both model-based and human annotations. Each example includes evaluations from 4 different frontier LLM models. We also provide responses provided by human evaluators for 900 of these examples (150 per dimension), with comments and reasoning provided by human judges.… See the full description on the dataset page: https://huggingface.co/datasets/anon34957/HumanAgencyEval_Human_Annotations.text10K<n<100K0 likes1 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.