datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vindr-cxr-testsetemu_edit_test_set
Dataset Card for the Emu Edit Test Set
Dataset Summary
To create a benchmark for image editing we first define seven different categories of potential image editing operations: background alteration (background), comprehensive image changes (global), style alteration (style), object removal (remove), object addition (add), localized modifications (local), and color/texture alterations (texture).
Then, we utilize the diverse set of input images from the MagicBrush… See the full description on the dataset page: https://huggingface.co/datasets/facebook/emu_edit_test_set.StableSR-TestSets
StableSR TestSets Card
These test sets are used associated with the StableSR, available here.
Data Details
Developed by: Jianyi Wang
Data type: Synthetic and real-world test sets for image super-resolution
License: S-Lab License 1.0
Data Description: The test sets are used to reproduce the metric results shown in Paper.
Resources for more information: GitHub Repository.
Cite as:
@InProceedings{wang2023exploiting,
author = {Wang, Jianyi and Yue, Zongsheng and… See the full description on the dataset page: https://huggingface.co/datasets/Iceclear/StableSR-TestSets.SpatialGen-Testset
SpatialGen Testset
This repository contains the test set for SPATIALGEN: Layout-guided 3D Indoor Scene Generation, a novel multi-view multi-modal diffusion model for generating realistic and semantically consistent 3D indoor scenes.
Project page | Paper | Code
We provide a test set of 48 preprocessed point clouds and their corresponding GT layouts, multi-view images are cropped from the high-resolution panoramic images.
Folder Structure
Outlines of the dataset files:… See the full description on the dataset page: https://huggingface.co/datasets/manycore-research/SpatialGen-Testset.indexed-open-image-v4-test-set
Dataset Card for "indexed-open-image-v4-test-set"
More Information needed
coco_test_set_pybboxes
COCO Test Set
This is a coco test set which is used for unit testing in pybboxes library.
MCABSA_testsetWorldRenderer-Testsetemu_edit_test_set_generations
Dataset Card for the Emu Edit Generations on Emu Edit Test Set
Dataset Summary
This dataset contains Emu Edit's generations on the Emu Edit test set. For more information please read our paper or visit our homepage.
Licensing Information
Licensed with CC-BY-NC 4.0 License available here.
Citation Information
@inproceedings{Sheynin2023EmuEP,
title={Emu Edit: Precise Image Editing via Recognition and Generation Tasks},
author={Shelly Sheynin and… See the full description on the dataset page: https://huggingface.co/datasets/facebook/emu_edit_test_set_generations.AIGCDetect_testsetAIGCDetect_testsetfrakturline-testset
Fraktur/Other Text-Line — Test Set
A balanced, held-out evaluation set of 2 000 scanned text-line images (1 000 per class) for the binary task of distinguishing Fraktur (blackletter / Gothic script) from other script (primarily Antiqua / Latin / Roman).
Developed for the Impresso digital humanities project.
Dataset Details
Property
Value
Task
Binary image classification
Classes
fraktur, other
Images per class
1 000
Total images
2 000
Image format
WebP… See the full description on the dataset page: https://huggingface.co/datasets/impresso-project/frakturline-testset.external_test_set_v1testset
Dataset Card for TreeOfLife-10M Captions
This dataset consists of generated captions, Wikipedia-derived descriptions and format examples for the TreeOfLife-10M. These captions were generated using InternVL3-38B based on biological contexts that help the model generate more accurate captions. It was used to train BioCAP, a CLIP-based model.
Dataset Details
This dataset is comprised of captions for the images in TreeOfLife-10M that were generated using InternVL3 38B.… See the full description on the dataset page: https://huggingface.co/datasets/ZihengZ/testset.winml-test-set
WinML Test Set
Dataset Summary
WinML Test Set is an evaluation‑only collection for validating model accuracy and stability on Windows ML / DirectML / ONNX Runtime pipelines. It aggregates several permissively‑licensed sources and harmonizes schema for reproducible, regression‑grade testing across backends and versions. Not intended for training.
Intended Use
Accuracy and regression benchmarking of Windows ML / DirectML / ONNX Runtime pipelines.… See the full description on the dataset page: https://huggingface.co/datasets/Futuremark/winml-test-set.Garments2Look-Test-Set-Results
Garments2Look: A Multi-Reference Dataset for High-Fidelity Outfit-Level Virtual Try-On with Clothing and Accessories
Project Page | Paper | Code
Garments2Look is a large-scale multimodal dataset for outfit-level Virtual Try-On (VTON), comprising 80,000 many-garments-to-one-look pairs across 40 major categories and over 300 fine-grained subcategories. Each pair includes an outfit with 3-12 reference garment images (averaging 4.48), a model image wearing the outfit, and detailed item… See the full description on the dataset page: https://huggingface.co/datasets/ArtmeScienceLab/Garments2Look-Test-Set-Results.FjordFish_Test_SetTest set for the FjordFish dataset https://zenodo.org/records/17950781 as seen in our publication (work in progress)
Overview of classes and annotations:
Class no.
Class
Images
Instances
all
200
830
0
ballan
5
5
1
cuckoo_female
2
2
2
cuckoo_male
4
4
3
corkwing_female
11
18
4
corkwing_male
35
36
5
goldsinny
83
226
6
rock_cook
5
6
7
labrids_unknown
34
81
8
cod59
106
11
pollack
7
8
13
saithe
4
10
14
whiting
7
11
15
gadines_unknown
37
56
16
spiny_dogfish
32… See the full description on the dataset page: https://huggingface.co/datasets/jdpoling/FjordFish_Test_Set.SpatialGen-Testset
SpatialGen Testset
This repository contains the test set for SPATIALGEN: Layout-guided 3D Indoor Scene Generation, a novel multi-view multi-modal diffusion model for generating realistic and semantically consistent 3D indoor scenes.
Project page | Paper | Code
We provide a test set of 48 preprocessed point clouds and their corresponding GT layouts, multi-view images are cropped from the high-resolution panoramic images.
Folder Structure
Outlines of the dataset… See the full description on the dataset page: https://huggingface.co/datasets/BenjaminChai579/SpatialGen-Testset.multimodal_online_RL_testsetpiper_flip-test-dragThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "custom_mpm_robot",
"total_episodes": 2,
"total_frames": 106,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Setsunainn/piper_flip-test-drag.GenAI-RealEstate-TestSet
🏙️ GenAI Real Estate Test Set (Track B)
Dataset for the MenaML Winter School 2026 Challenge.
📊 Dataset Structure
This dataset contains 1,000 images split evenly between:
Authentic: Real estate photography from the Places365 dataset.
Manipulated: Synthetically generated deepfake artifacts (Inpainting, Diffusion Noise, GAN Grids).
🕵️ How to Use
This dataset is designed for testing forensic detection models.
test_setpiper_flip-test-turnThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "custom_mpm_robot",
"total_episodes": 2,
"total_frames": 272,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Setsunainn/piper_flip-test-turn.DocGenome-Testset-DocQA
DocGenome-TestSet-DocQA
Firstly, you need to download the test dataset here.
Then, unzip the DocGenome-Testset-DocQA.zip as follows:
DocGenome-Testset-DocQA
├── testset
│ ├── xxx
├── eval_tools
│ ├── eval_open_docqa_gpt.py
│ ├── eval_normal_docqa.py
├── qa_info
│ ├── docgenome_testset_multiqa.json
│ ├── docgenome_testset_singleqa.json
│ ├── docgenome_testset_normalqa.jsonl
├── example
│ ├── internvl_open_docqa_test.py
│ ├── internvl_normal_docqa_test.py
├──… See the full description on the dataset page: https://huggingface.co/datasets/U4R/DocGenome-Testset-DocQA.piper_flip-test-v0-turnThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "custom_mpm_robot",
"total_episodes": 2,
"total_frames": 272,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Setsunainn/piper_flip-test-v0-turn.clevr_test_extent_settinggemini-finetune-test-setpiper_flip-test-pullThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "custom_mpm_robot",
"total_episodes": 2,
"total_frames": 806,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Setsunainn/piper_flip-test-pull.my_robot_set_testThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 2,
"total_frames": 557,
"total_tasks": 1,
"total_videos": 4,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/gaozj/my_robot_set_test.ChartQA_testset
