datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
OR-Space
OR-Space
A full-lifecycle workspace benchmark for industrial optimization agents.
OR-Space evaluates whether language-model agents can work reliably with
operations research problems represented as executable, multi-file workspaces.
Rather than presenting a self-contained mathematical prompt, each task
distributes evidence across business requirements, structured data, source
code, execution logs, and solver records.
The benchmark contains 100 optimization topologies. Each… See the full description on the dataset page: https://huggingface.co/datasets/Chenyu-Zhou/OR-Space.CDial-BiasOfficial release of CDial-Bias dataset.
Notation:
Before downloading the dataset, please be aware that: The CDial-Bias Dataset is released for research purpose only and other usages require further permission. Please ensure the usage contributes to improving the safety and fairness of AI technologies. No malicious usage is allowed.
Paper:
https://aclanthology.org/2022.findings-emnlp.262/
Github Repo:
https://github.com/para-zhou/CDial-Bias
Leaderboard:… See the full description on the dataset page: https://huggingface.co/datasets/para-zhou/CDial-Bias.BaboonLand
Dataset Card for BaboonLand Dataset: Tracking Primates in the Wild and Automating Behaviour Recognition from Drone Videos
Dataset Summary
BaboonLand is an aerial drone video dataset of wild olive baboons (Papio anubis) collected over 21 consecutive days in Laikipia (Mpala Research Centre), Kenya, following three troops during morning and evening movements to and from sleeping sites. The dataset contains UAV footage across diverse environments (e.g., sleeping tree, river… See the full description on the dataset page: https://huggingface.co/datasets/zhoubingyu/BaboonLand.KABR
Dataset Card for KABR: In-Situ Dataset for Kenyan Animal Behavior Recognition from Drone Videos
Dataset Summary
We present a novel high-quality dataset for animal behavior recognition from drone videos.
The dataset is focused on Kenyan wildlife and contains behaviors of giraffes, plains zebras, and Grevy's zebras.
The dataset consists of more than 10 hours of annotated videos, and it includes eight different classes, encompassing seven types of animal behavior and an… See the full description on the dataset page: https://huggingface.co/datasets/zhoubingyu/KABR.MMVMBench_VQAscrape-content-dataset-v1
Scrape Content Dataset v1
A human-curated benchmark dataset for evaluating web scraping engines on content quality.
Overview
This dataset contains 1,000 web pages with human-annotated ground truth for evaluating how well web scraping engines capture core content while avoiding noise (navigation, ads, footers, etc.). The dataset was created in 2025-10-21 and may become outdated over time.
Dataset Structure
CSV format with columns:
id: Sequential… See the full description on the dataset page: https://huggingface.co/datasets/zhoubinghong/scrape-content-dataset-v1.A-sharesGameLabel-10kGameLabel-10k Dataset Card
This dataset contains was created in collaboration with the game developers of Armchair Commander. It contains 9800 human preferences over pairs of Flux-Schnell generated images, with over 6800 unique prompts. All labels were crowdsourced from Armchair Commander players.
Usage Example
from datasets import load_dataset
from PIL import Image
import base64
from io import BytesIO
dataset = load_dataset("Jonathan-Zhou/GameLabel-10k")
# For some reason, when using… See the full description on the dataset page: https://huggingface.co/datasets/Jonathan-Zhou/GameLabel-10k.OmniShow_example_dataset
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
Donghao Zhou1,*, Guisheng Liu2,*, Hao Yang2, Jiatong Li2,†, Jingyu Lin3, Xiaohu Huang4,
Yichen Liu2, Xin Gao2, Cunjian Chen3, Shilei Wen2,§, Chi-Wing Fu1, Pheng-Ann Heng1,§
1The Chinese University of Hong Kong, 2ByteDance, 3Monash University, 4The University of Hong Kong
*Equal contribution, †Project lead, §Corresponding author
🌍 Useful Links
Project Page:… See the full description on the dataset page: https://huggingface.co/datasets/donghao-zhou/OmniShow_example_dataset.ZhouhcReliabilityBench
Dataset Card for ReliabilityBench
Dataset Summary
ReliabilityBench is a benchmark with multiple datasets across five domains, introduced in the paper: Larger and More Instructable Language Models Become Less Reliable. Lexin Zhou, Wout Schellaert, Fernando Martı́nez-Plumed, Yael Moros-Daval, Cèsar Ferri, and José Hernández-Orallo.
The five domains correspond to: simple numeracy (‘addition’), vocabulary reshuffle (‘anagram’), geographical knowledge (‘locality’), basic and… See the full description on the dataset page: https://huggingface.co/datasets/lexin-zhou/ReliabilityBench.Chinese-Student-English-Essay
Dataset Card for Chinese Student English Essay (CSEE) Dataset
Dataset Summary
The Chinese Student English Essay (CSEE) dataset is designed for Automated Essay Scoring (AES) tasks. It consists of 13,270 English essays written by high school students in Beijing, who are English as a Second Language (ESL) learners. These essays were collected from two final exams and correspond to two writing prompts.
Each essay is evaluated across three key dimensions by experienced… See the full description on the dataset page: https://huggingface.co/datasets/zhongshiting/Chinese-Student-English-Essay.Contenttlile-contentDemenCare
DemenCare
Question–answer items (qa subset) and human ratings (scores subset). Join on QAID.
Subsets
Config
File
Description
qa
DemenCare_QA.csv
QAID, StudyNumber, QuestionID, Source, Question, Answer
scores
DemenCare_scores.csv
QAID, RaterID, RaterGroup, rubric dimensions
License
See dataset card on the Hub for license and citation.
zhorzh-simenon-megre-i-chalavek-na-lautsy
Мэгрэ і чалавек на лаўцы
Metadata
Author: Жорж Сімэнон
Title: Мэгрэ і чалавек на лаўцы
Narrator:
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum split… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/zhorzh-simenon-megre-i-chalavek-na-lautsy.RobertFrostDEMIMathAnalysis
