datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
X-EGO-CS
X-Ego-CS
Ten players. One match. Ten simultaneous first-person recordings, each paired
with a 64 Hz stream of that player's exact keyboard, mouse and view-angle
inputs — all on a common, measured clock.
Paper · Paper code · Collection pipeline
Cross-Ego Demo (Pistol Round)
Your browser cannot play this video —
download it instead.
All ten players' points of view, from the same pistol round, on one clock.
Note: this demo concatenates the ten streams… See the full description on the dataset page: https://huggingface.co/datasets/wangyz1999/X-EGO-CS.xstory_cloze
Dataset Card for XStoryCloze
Dataset Summary
XStoryCloze consists of the professionally translated version of the English StoryCloze dataset (Spring 2016 version) to 10 non-English languages. This dataset is released by Meta AI.
Supported Tasks and Leaderboards
commonsense reasoning
Languages
en, ru, zh (Simplified), es (Latin America), ar, hi, id, te, sw, eu, my.
Dataset Structure
Data Instances
Size of downloaded dataset… See the full description on the dataset page: https://huggingface.co/datasets/juletxara/xstory_cloze.ptb-xl-processedeccv2026-cad-challenge-data
ECCV 2026 CAD Challenge Data
This challenge is part of the workshop The Path to Manufacturing: Evolving
3D Generation to Intelligent Computer-Aided Design.
Workshop homepage: https://3dgen-cad-workshop.github.io/
Challenge submission Space: https://huggingface.co/spaces/jingwei-xu-00/eccv2026-cad-challenge
Dataset rendering and preparation code (only .step files are required): https://github.com/DavidXu-JJ/eccv2026-cad-challenge-data-render
This repository contains the public… See the full description on the dataset page: https://huggingface.co/datasets/jingwei-xu-00/eccv2026-cad-challenge-data.HR-VILAGE-3K3M
HR-VILAGE-3K3M: Human Respiratory Viral Immunization Longitudinal Gene Expression
This repository provides the HR-VILAGE-3K3M dataset, a curated collection of human longitudinal gene expression profiles, antibody measurements, and aligned metadata from respiratory viral immunization and infection studies. The dataset includes baseline transcriptomic profiles and covers diverse exposure types (vaccination, inoculation, and mixed exposure). HR-VILAGE-3K3M is designed as a… See the full description on the dataset page: https://huggingface.co/datasets/xuejun72/HR-VILAGE-3K3M.XSTest
XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models
XSTest is a test suite designed to identify exaggerated safety / false refusal in Large Language Models (LLMs).
It comprises 250 safe prompts across 10 different prompt types, along with 200 unsafe prompts as contrasts.
The test suite aims to evaluate how well LLMs balance being helpful with being harmless by testing if they unnecessarily refuse to answer safe prompts that superficially… See the full description on the dataset page: https://huggingface.co/datasets/Paul/XSTest.Objaverse-XL-Rigged-Animated
Objaverse-XL Rigged & Animated Subset
Every asset here carries both a skeleton and at least one animation clip, selected from
Objaverse / Objaverse-XL. Rigs range from 3 to 344 joints and
span characters as well as articulated rigid objects.
Objaverse-XL indexes over 10 million objects, but only a small fraction carry a usable rig and
motion on it. This subset isolates that fraction: every file was checked to contain at least one
skin with joints and at least one animation clip… See the full description on the dataset page: https://huggingface.co/datasets/Linzhan/Objaverse-XL-Rigged-Animated.publichealth-qa
Usage
import datasets
langs = ['arabic', 'chinese', 'english', 'french', 'korean', 'russian', 'spanish', 'vietnamese']
data = datasets.load_dataset('xhluca/publichealth-qa', split='test', name=langs[0])
About
This dataset contains question and answer pairs sourced from Q&A pages and FAQs from CDC and WHO pertaining to COVID-19. They were produced and collected between 2019-12 and 2020-04. They were originally published as an aggregated Kaggle dataset.… See the full description on the dataset page: https://huggingface.co/datasets/xhluca/publichealth-qa.Frappe_x1
Frappe_x1
Dataset description:
The Frappe dataset contains a context-aware app usage log, which comprises 96203 entries by 957 users for 4082 apps used in various contexts. It has 10 feature fields including user_id, item_id, daytime, weekday, isweekend, homework, cost, weather, country, city. The target value indicates whether the user has used the app under the context. Following the AFN work, we randomly split the data into 7:2:1 as the training set, validation set, and test set… See the full description on the dataset page: https://huggingface.co/datasets/reczoo/Frappe_x1.CUHK-X_Small_Model_Track
CUHK-X — Small Model Track
Multimodal human action recognition (classification).
Given a multimodal clip, predict its action class (action_id, 0–39, 40 classes).
Repository layout
.
├── Training/
│ ├── class_mapping.csv # action_id <-> action_name (40 classes)
│ └── data/
│ └── HAR.z01 … HAR.z08 + HAR.zip # multi-volume zip
│ → HAR/data/<modality>/<action>/<user>/<trial>/<files>
└── Testing/
├── data/
│ └──… See the full description on the dataset page: https://huggingface.co/datasets/Kevin-Pal/CUHK-X_Small_Model_Track.w2t-llm-arc-easy-lora
W2T Llm Arc Easy Lora
This repository contains artifacts for the W2T paper:
Paper: W2T: LoRA Weights Already Know What They Can Do
Repo: Weight2Token
Summary
ARC-Easy LoRA checkpoints and prepared metadata used for performance prediction.
Source Status
Storage location: local
Verification status: confirmed
Files
See manifest.json for the exact local or remote source paths used to prepare this release.
Citation… See the full description on the dataset page: https://huggingface.co/datasets/Xiaolong-Han/w2t-llm-arc-easy-lora.ICPC_Data
ICPC World Finals — a discriminative subset, with model traces
24 ICPC World Finals problems (2021–2025), together with the full transcripts of an
LLM attempting each of them three times under simulated contest rules.
Selection
The model
Every run in this dataset comes from:
nvidia/Nemotron-Cascade-2-30B-A3B
The partitions
Every one of the 53 problems was run 3 times (seeds 1, 2, 3). Each problem was then
placed by its pass rate and… See the full description on the dataset page: https://huggingface.co/datasets/xupy21/ICPC_Data.pxr-structure-pose-pool
PXR Structure Challenge — Full Multi-Model Pose Pool (184 ligands)
Every protein–ligand pose generated during the OpenADMET PXR (pregnane X receptor / NR1I2)
structure-prediction challenge, released openly with per-pose labels so the community can
reuse the compute already spent — and, we hope, crack the problem this data makes visible.
What's here
poses/<model>/<SID>.pdb — one best pose per (model, ligand). Protein chain A + ligand
(resname LIG). 15 models, up… See the full description on the dataset page: https://huggingface.co/datasets/xX-its-amit-Xx/pxr-structure-pose-pool.DNA_Gen
Citation
Please cite our work using the bibtex below:
BibTeX:
@article{su2025language,
title={Language Models for Controllable DNA Sequence Design},
author={Su, Xingyu and Li, Xiner and Lin, Yuchao and Xie, Ziqian and Zhi, Degui and Ji, Shuiwang},
journal={arXiv preprint arXiv:2507.19523},
year={2025}
}
DeepSearch
xbench-evals
🌐 Website | 📄 Paper | 🤗 Dataset
Evergreen, contamination-free, real-world, domain-specific AI evaluation framework
xbench is more than just a scoreboard — it's a new evaluation framework with two complementary tracks, designed to measure both the intelligence frontier and real-world utility of AI systems:
AGI Tracking: Measures core model capabilities like reasoning, tool-use, and memory
Profession Aligned: A new class of evals grounded in workflows, environments… See the full description on the dataset page: https://huggingface.co/datasets/xbench/DeepSearch.xauusd
Cleaned XAUUSD Dataset
Dataset Description
This dataset contains cleaned and preprocessed minute-level historical price data for the XAU/USD (Gold vs. US Dollar) pair. The data spans from November 1, 2011, to January 3, 2024, and includes the following columns:
open: The opening price of the minute.
high: The highest price during the minute.
low: The lowest price during the minute.
close: The closing price of the minute.
tickvol: The number of price changes (ticks)… See the full description on the dataset page: https://huggingface.co/datasets/Pcitycrypto/xauusd.DAAD-XAbout Dataset
The DAADX Dataset is derived from DAAD dataset (https://cvit.iiit.ac.in/research/projects/cvit-projects/daad#dataset), which contains all the captured videos for the Driver Intention Prediction task. We are introducing the first video based explanations dataset for driver intention prediction task. This will be help in further the research interms of making an explainable Autonomous Driving or ADAS System.
DAAD-X contains explanations for each maneuver instance, these… See the full description on the dataset page: https://huggingface.co/datasets/Skyrmion/DAAD-X.linkedin-job-postingsVStar_Bench
V* Benchmark
refactor V* Benchmark to add support for VLMEvalKit.
Benchmark Information
Number of questions: 191
Question type: MCQ (multiple choice question)
Question format: image + text
Reference
VLMEvalKit
V* Benchmark
MMVP
xlwic_wn
Multilingual Word-in-Context (WordNet)
Refer to the documentation and paper for more information.
DeepSearch-2510
xbench-evals
🌐 Website | 📄 Paper | 🤗 Dataset
Evergreen, contamination-free, real-world, domain-specific AI evaluation framework
xbench is more than just a scoreboard — it's a new evaluation framework with two complementary tracks, designed to measure both the intelligence frontier and real-world utility of AI systems:
AGI Tracking: Measures core model capabilities like reasoning, tool-use, and memory
Profession Aligned: A new class of evals grounded in workflows, environments… See the full description on the dataset page: https://huggingface.co/datasets/xbench/DeepSearch-2510.x402-market-pulse
x402 Market Pulse
A small, growing time series of how much USDC actually settles through x402 pay-per-call APIs on Base,
measured from on-chain Transfer logs rather than from provider-reported call counts. One row per
measurement; every row is a trailing 24-hour window ending at measured_at_utc.
Status: bootstrapping. 74 rows spanning 2026-09-22T13:34Z to 2026-09-25T18:00Z —
less than one independent day of data so far. Rows are kept to at most one per clock hour and a new one… See the full description on the dataset page: https://huggingface.co/datasets/nimapro1381/x402-market-pulse.Criteo_x1
Criteo_x1
Dataset description:
The Criteo dataset is a widely-used benchmark dataset for CTR prediction, which contains about one week of click-through data for display advertising. It has 13 numerical feature fields and 26 categorical feature fields. Following the AFN work, we randomly split the data into 7:2:1* as the training set, validation set, and test set, respectively.
The dataset statistics are summarized as follows:
Dataset Split
Total
#Train
#Validation
#Test… See the full description on the dataset page: https://huggingface.co/datasets/reczoo/Criteo_x1.3D-RAD
[ 🎯 NeurIPS 2025 ] 3D-RAD 🩻: A Comprehensive 3D Radiology Med-VQA Dataset with Multi-Temporal Analysis and Diverse Diagnostic Tasks
📢 News
What's New in This Update 🚀
2025.10.23: 🔥 Updated the latest version of the paper!
2025.09.19: 🔥 Paper accepted to NeurIPS 2025! 🎯
2025.05.16: 🔥 Set up the repository and committed the dataset!
🔍 Overview
💡 In this repository, we present the dataset for "3D-RAD: A… See the full description on the dataset page: https://huggingface.co/datasets/Tang-xiaoxiao/3D-RAD.patientcare
Patient Care Activity Recognition Dataset
Dataset Summary
A multi-class video action recognition dataset covering 23 clinically relevant patient care activities recorded in hospital-like environments. The dataset includes both real-life recordings and synthetically generated videos.
The dataset is intended to support research in automated patient monitoring and clinical activity recognition using computer vision.
Supported Tasks
Video… See the full description on the dataset page: https://huggingface.co/datasets/xyz1901901/patientcare.xai-studies
xai-studies
xai = explainable AI. Offline mirror of lyffseba/xai.
GitHub
lyffseba/xai
portal
spaces/lyffseba/xai
use
python3 studies/run.py test
No pip. No network.
catalog
catalog/models.csv
id
name
status
total
active
ctx
experts
inkling
Inkling
weights_public
975B
41B
1048576
6/256+2 shared
kimi-k3
Kimi K3
api_live_weights_pending
2.8T
1048576
16/896
laguna-s-2.1
Laguna S 2.1
weights_public
118B
8B
1048576… See the full description on the dataset page: https://huggingface.co/datasets/lyffseba/xai-studies.rednote-xiaohongshu-notes
RedNote (Xiaohongshu) Notes with Engagement and Save Rates
10,427 posts from RedNote (小红书 / Xiaohongshu), across 8 content verticals, with full Chinese post text, four separate engagement metrics, and province-level geography.
Most social datasets give you text and a like count. RedNote separates saves from likes, and that single distinction turns out to measure something the like count cannot.
The finding this dataset exists for
A post can be useful or it can be… See the full description on the dataset page: https://huggingface.co/datasets/FreshCrawl/rednote-xiaohongshu-notes.xnli-eu
Dataset Card for XNLIeu
XNLIeu is an extension of XNLI translated from English to Basque. It has been designed as a cross-lingual dataset for the Natural Language Inference task, a text-classification task that consists on classifying pairs of sentences, a premise and a hypothesis, according to their semantic relation out of three possible labels: entailment, contradiction and neutral.
Dataset Details
Dataset Description
XNLI is a popular Natural… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/xnli-eu.extractive_qa_question_answering_hr
Dataset Card
HR-Multiwoz is a fully-labeled dataset of 5980 extractive qa spanning 10 HR domains to evaluate LLM Agent. It is the first labeled open-sourced conversation dataset in the HR domain for NLP research.
Please refer to HR-MultiWOZ: A Task Oriented Dialogue (TOD) Dataset for HR LLM Agent for details about the dataset construction.
Dataset Sources
Repository: xwjzds/extractive_qa_question_answering_hr
Paper: HR-MultiWOZ: A Task Oriented Dialogue (TOD)… See the full description on the dataset page: https://huggingface.co/datasets/xwjzds/extractive_qa_question_answering_hr.x402-price-index
x402 Price Index — a census of the machine-payable web
Most directories of x402 services list what merchants claim. This dataset records
what their endpoints actually answer when probed: the HTTP 402 challenge they
return, the price inside it, the chain and asset they want, and whether they respond
at all.
It is built by Animica from an independent prober that walks
every x402 endpoint it can discover and reads each merchant's own accepts[] block.
No merchant self-reporting is… See the full description on the dataset page: https://huggingface.co/datasets/animicaorg/x402-price-index.
