CoolFace
18 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Chenyu-Zhou /OR-Space OR-Space A full-lifecycle workspace benchmark for industrial optimization agents. OR-Space evaluates whether language-model agents can work reliably with operations research problems represented as executable, multi-file workspaces. Rather than presenting a self-contained mathematical prompt, each task distributes evidence across business requirements, structured data, source code, execution logs, and solver records. The benchmark contains 100 optimization topologies. Each… See the full description on the dataset page: https://huggingface.co/datasets/Chenyu-Zhou/OR-Space.textquestion-answeringn<1K4 likes634 downloads2mo agoHugging Face02para-zhou /CDial-BiasOfficial release of CDial-Bias dataset. Notation: Before downloading the dataset, please be aware that: The CDial-Bias Dataset is released for research purpose only and other usages require further permission. Please ensure the usage contributes to improving the safety and fairness of AI technologies. No malicious usage is allowed. Paper: https://aclanthology.org/2022.findings-emnlp.262/ Github Repo: https://github.com/para-zhou/CDial-Bias Leaderboard:… See the full description on the dataset page: https://huggingface.co/datasets/para-zhou/CDial-Bias.tabulartext-classification10K<n<100K8 likes544 downloads2y agoHugging Face03zhoubingyu /BaboonLand Dataset Card for BaboonLand Dataset: Tracking Primates in the Wild and Automating Behaviour Recognition from Drone Videos Dataset Summary BaboonLand is an aerial drone video dataset of wild olive baboons (Papio anubis) collected over 21 consecutive days in Laikipia (Mpala Research Centre), Kenya, following three troops during morning and evening movements to and from sleeping sites. The dataset contains UAV footage across diverse environments (e.g., sleeping tree, river… See the full description on the dataset page: https://huggingface.co/datasets/zhoubingyu/BaboonLand.textvideo-classification1M<n<10M0 likes431 downloads6mo agoHugging Face04zhoubingyu /KABR Dataset Card for KABR: In-Situ Dataset for Kenyan Animal Behavior Recognition from Drone Videos Dataset Summary We present a novel high-quality dataset for animal behavior recognition from drone videos. The dataset is focused on Kenyan wildlife and contains behaviors of giraffes, plains zebras, and Grevy's zebras. The dataset consists of more than 10 hours of annotated videos, and it includes eight different classes, encompassing seven types of animal behavior and an… See the full description on the dataset page: https://huggingface.co/datasets/zhoubingyu/KABR.textvideo-classification1M<n<10M0 likes49 downloads5mo agoHugging Face05zhouyik /MMVMBench_VQAtext1K<n<10K0 likes46 downloads1y agoHugging Face06zhoubinghong /scrape-content-dataset-v1 Scrape Content Dataset v1 A human-curated benchmark dataset for evaluating web scraping engines on content quality. Overview This dataset contains 1,000 web pages with human-annotated ground truth for evaluating how well web scraping engines capture core content while avoiding noise (navigation, ads, footers, etc.). The dataset was created in 2025-10-21 and may become outdated over time. Dataset Structure CSV format with columns: id: Sequential… See the full description on the dataset page: https://huggingface.co/datasets/zhoubinghong/scrape-content-dataset-v1.text1K<n<10K0 likes45 downloads7d agoHugging Face07Jinsong-Zhou /A-sharesgatedtabular100M<n<1B5 likes44 downloads2y agoHugging Face08Jonathan-Zhou /GameLabel-10kGameLabel-10k Dataset Card This dataset contains was created in collaboration with the game developers of Armchair Commander. It contains 9800 human preferences over pairs of Flux-Schnell generated images, with over 6800 unique prompts. All labels were crowdsourced from Armchair Commander players. Usage Example from datasets import load_dataset from PIL import Image import base64 from io import BytesIO dataset = load_dataset("Jonathan-Zhou/GameLabel-10k") # For some reason, when using… See the full description on the dataset page: https://huggingface.co/datasets/Jonathan-Zhou/GameLabel-10k.tabular1K<n<10K0 likes36 downloads2y agoHugging Face09donghao-zhou /OmniShow_example_dataset OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Donghao Zhou1,*, Guisheng Liu2,*, Hao Yang2, Jiatong Li2,†, Jingyu Lin3, Xiaohu Huang4, Yichen Liu2, Xin Gao2, Cunjian Chen3, Shilei Wen2,§, Chi-Wing Fu1, Pheng-Ann Heng1,§ 1The Chinese University of Hong Kong, 2ByteDance, 3Monash University, 4The University of Hong Kong *Equal contribution, †Project lead, §Corresponding author 🌍 Useful Links Project Page:… See the full description on the dataset page: https://huggingface.co/datasets/donghao-zhou/OmniShow_example_dataset.audion<1K0 likes20 downloads5mo agoHugging Face10Littleoops /Zhouhctext10K<n<100K0 likes12 downloads2y agoHugging Face11lexin-zhou /ReliabilityBenchgated Dataset Card for ReliabilityBench Dataset Summary ReliabilityBench is a benchmark with multiple datasets across five domains, introduced in the paper: Larger and More Instructable Language Models Become Less Reliable. Lexin Zhou, Wout Schellaert, Fernando Martı́nez-Plumed, Yael Moros-Daval, Cèsar Ferri, and José Hernández-Orallo. The five domains correspond to: simple numeracy (‘addition’), vocabulary reshuffle (‘anagram’), geographical knowledge (‘locality’), basic and… See the full description on the dataset page: https://huggingface.co/datasets/lexin-zhou/ReliabilityBench.tabular100K<n<1M8 likes10 downloads2y agoHugging Face12zhongshiting /Chinese-Student-English-Essay Dataset Card for Chinese Student English Essay (CSEE) Dataset Dataset Summary The Chinese Student English Essay (CSEE) dataset is designed for Automated Essay Scoring (AES) tasks. It consists of 13,270 English essays written by high school students in Beijing, who are English as a Second Language (ESL) learners. These essays were collected from two final exams and correspond to two writing prompts. Each essay is evaluated across three key dimensions by experienced… See the full description on the dataset page: https://huggingface.co/datasets/zhongshiting/Chinese-Student-English-Essay.tabulartext-classification10K<n<100K0 likes10 downloads5mo agoHugging Face13zhouzhouJF /Contenttext1K<n<10K0 likes6 downloads1y agoHugging Face14zhouzhouJF /tlile-contenttext100K<n<1M0 likes3 downloads1y agoHugging Face15ZhouLabAI /DemenCare DemenCare Question–answer items (qa subset) and human ratings (scores subset). Join on QAID. Subsets Config File Description qa DemenCare_QA.csv QAID, StudyNumber, QuestionID, Source, Question, Answer scores DemenCare_scores.csv QAID, RaterID, RaterGroup, rubric dimensions License See dataset card on the Hub for license and citation. tabularn<1K0 likes2 downloads5mo agoHugging Face16archivartaunik /zhorzh-simenon-megre-i-chalavek-na-lautsy Мэгрэ і чалавек на лаўцы Metadata Author: Жорж Сімэнон Title: Мэгрэ і чалавек на лаўцы Narrator: Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/zhorzh-simenon-megre-i-chalavek-na-lautsy.audion<1K0 likes2 downloads4mo agoHugging Face17Zhoucai /RobertFrosttext1K<n<10K0 likes1 downloads3y agoHugging Face18zhouliang /DEMIMathAnalysisgatedtextn<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.