datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DatasetWithCapitalLetterscyp-challenge-train-test
CYP Challenge Train/Test Dataset
A high-quality experimental dataset for predicting inhibition of the major drug-metabolizing Cytochrome P450 enzymes (CYP1A2, CYP2C9, CYP2D6, CYP3A4), released as part of the OpenADMET CYP Inhibition Blind Challenge.
Blog post: Announcing OpenADMET’s CYP inhibition blind challenge
Challenge Space: OpenADMET CYP Inhibition Blind Challenge
Challenge period: August 17, 2026 - November 3, 2026
Produced by: OpenADMET
CHANGELOG
Updated… See the full description on the dataset page: https://huggingface.co/datasets/openadmet/cyp-challenge-train-test.protein_data_testsplit 1, 2 -> for sequences
split 3, 4 -> for residues
SpatialLM-Testset
SpatialLM Testset
Project page | Paper | Code
We provide a test set of 107 preprocessed point clouds and their corresponding GT layouts, point clouds are reconstructed from RGB videos using MASt3R-SLAM. SpatialLM-Testset is quite challenging compared to prior clean RGBD scan datasets due to the noises and occlusions in the point clouds reconstructed from monocular RGB videos.
Folder Structure
Outlines of the dataset files:… See the full description on the dataset page: https://huggingface.co/datasets/manycore-research/SpatialLM-Testset.pxr-challenge-train-test
PXR Challenge Train/Test Dataset
A high-quality experimental dataset for predicting human Pregnane-X Receptor (PXR) induction, comprising over 11,000 compounds screened using a high-fidelity in-house assay. This is the largest publicly available PXR activity dataset, released as part of the OpenADMET PXR Induction Blind Challenge.
Blog post: Announcing the Next OpenADMET Blind Challenge: Predicting PXR Induction
Challenge Space: openadmet/pxr-challenge
Challenge period: April 1… See the full description on the dataset page: https://huggingface.co/datasets/openadmet/pxr-challenge-train-test.bias-test-gpt-sentences
Dataset Card for "BiasTestGPT: Generated Test Sentences"
Dataset of sentences for bias testing in open-sourced Pretrained Language Models generated using ChatGPT and other generative Language Models.
This dataset is used and actively populated by the BiasTestGPT HuggingFace Tool.
BiasTestGPT HuggingFace Tool
Dataset with Bias Specifications
Project Landing Page
Dataset Structure
The dataset is structured as a set of CSV files with names corresponding to the social… See the full description on the dataset page: https://huggingface.co/datasets/AnimaLab/bias-test-gpt-sentences.contextual_testCheck out the paper.
vindr-cxr-testsetcasp14-casp15-cameo-test-proteinsprotein_data_test_2openadmet-expansionrx-challenge-test-data-blinded
OpenADMET-ExpansionRx Challenge blinded test dataset
This dataset contains real-work ADMET data from a recently prosecuted series of drug discovery campaigns by Expansion Therapeutics on RNA mediated diseases. While optimising candidate molecules for their preclinical programs Expansion collected a variety of ADMET data for off-targets and properties of interest in the traditional game of “whack-a-mole” familiar to all drug hunters. Now, they’ve made the bold and generous decision… See the full description on the dataset page: https://huggingface.co/datasets/openadmet/openadmet-expansionrx-challenge-test-data-blinded.videophy2_testProject: https://github.com/Hritikbansal/videophy/tree/main/VIDEOPHY2
caption: original prompt in the dataset
video_url: generated video (using original prompt or upsampled caption, depending on the video model)
sa: semantic adherence score (1-5) from human evaluation
pc: physical commonsense score (1-5) from human evaluation
joint: computed as sa >= 4, pc >= 4
physics_rules_followed: list of physics rules followed in the video as judged by human annotators (1)
physics_rules_unfollowed: list… See the full description on the dataset page: https://huggingface.co/datasets/videophysics/videophy2_test.test2testset_popqavideophy_test_publicWe have uploaded the videos at: https://huggingface.co/videophysics/videophy-test-videos/tree/main
For more details, please visit:
project github: https://github.com/Hritikbansal/videophy
project website: https://videophy.github.io/
test-sciencetestset_piqaUWM-test
UWM-test
A held-out test set sampled from four RoboMIND 2.0 robot-manipulation subsets
(Franka, Tianyi, Tienkung, UR5). For every task that belongs to the test split
(is_training = 0), 25 episodes are sampled. UR5 uses a pre-existing
8 × 25 subtest selection.
Total: 600 episodes, 24 tasks, ~288 GiB.
Directory layout
UWM-test/
├── uwm_test_index.csv # master index (paths point into this folder)
├── README.md
├── franka/<task>/<episode_id>/
│ ├──… See the full description on the dataset page: https://huggingface.co/datasets/zmliu1122/UWM-test.test_tsv_datasetargilla-invalid-rowstest_many_filesbias-test-gpt-sentencesFARM_training_test
FARM Aerial Radio Map (ARM) Dataset
Paper:
FARM: Foundational Aerial Radio Map for Intelligent Low-Altitude Networking (https://arxiv.org/abs/2604.17362)
Overview
This repository releases the constructed ARM datasets based on ARM-Omni for FARM training, in-domain evaluation (D1-D10), and zero-shot evaluation (P1, F1, and A1). The dataset coverage is summarized below:
Dataset
Frequencies (GHz)
Max Rx Height (m)
Beamwidths
Map Grid Size
Volume
D1
2.1… See the full description on the dataset page: https://huggingface.co/datasets/jliang097/FARM_training_test.patents_claims_1.5m_traim_test2025_Virtual_Cell_Challenge_Test_Datatestset_mmludialogsum-test
Dataset Card for DIALOGSum Corpus
Dataset Description
Links
Homepage: https://aclanthology.org/2021.findings-acl.449
Repository: https://github.com/cylnlp/dialogsum
Paper: https://aclanthology.org/2021.findings-acl.449
Point of Contact: https://huggingface.co/knkarthick
Dataset Summary
DialogSum is a large-scale dialogue summarization dataset, consisting of 13,460 (Plus 100 holdout data for topic generation) dialogues with corresponding… See the full description on the dataset page: https://huggingface.co/datasets/neil-code/dialogsum-test.testset_hellaswagjava_unit_testtest-translation-dataset
