datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
PhysicalAI-WorldModel-Synthetic-Physical-Interaction-Scenes
PhysicalAI-WorldModel-Synthetic-Physical-Interaction-Scenes Dataset Card
Dataset Description
PhysicalAI-WorldModel-Synthetic-Physical-Interaction-Scenes is a large-scale synthetic dataset of physically-simulated multi-object interaction scenes, generated using NVIDIA Isaac Sim and the PhysX physics engine. It is designed to train and evaluate AI models on physical reasoning, rigid body dynamics, optical flow, depth estimation, and scene understanding.
Each clip… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/PhysicalAI-WorldModel-Synthetic-Physical-Interaction-Scenes.synthetic_vocal_burstsThis repository contains the vocal bursts like giggling, laughter, shouting, crying, etc. from the following repository.
https://huggingface.co/datasets/sleeping-ai/Vocal-burst
We captioned them using Gemini Flash Audio 2.0. This dataset contains, this dataset contains ~ 365,000 vocal bursts from all kinds of categories.
It might be helpful for pre-training audio text foundation models to generate and understand all kinds of nuances in vocal bursts.
synthetic-cyrillic-largesynthetic-derm-1M-trainTransNormal-Synthetic
TransNormal-Synthetic Dataset
Physics-based synthetic dataset for transparent object normal estimation.
This dataset accompanies the paper:
TransNormal: Dense Visual Semantics for Diffusion-based Transparent Object Normal Estimation
Mingwei Li, Hehe Fan, Yi Yang
arXiv:2602.00839 | Project Page | Code
Dataset Description
TransNormal-Synthetic is a physics-based rendered dataset featuring transparent laboratory equipment (beakers, flasks, test tubes, etc.) with… See the full description on the dataset page: https://huggingface.co/datasets/Longxiang-ai/TransNormal-Synthetic.more-synthetic-vocalbursts-raw
More Synthetic Vocal Bursts (Raw)
Synthetic vocal burst audio samples generated from a taxonomy of 202 vocal burst types across multiple text-to-audio and TTS models. Each sample is a short (3–10 second) non-speech vocalization — laughs, cries, gasps, sighs, growls, etc. — generated from text prompts describing the burst type, gender, and age group.
Models Used
Model
Type
Samples
Sample Rate
Notes
DramaBox (ResembleAI/Dramabox)
TTS DiT
2000
44.1 kHz… See the full description on the dataset page: https://huggingface.co/datasets/laion/more-synthetic-vocalbursts-raw.synthetic-dataset-1m-dalle3-high-quality-captions
Dataset Card for Dalle3 1 Million+ High Quality Captions
Alt name: Human Preference Synthetic Dataset
Example grids for landscapes, cats, creatures, and fantasy are also available.
Description:
This dataset comprises of AI-generated images sourced from various websites and individuals, primarily focusing on Dalle 3 content, along with contributions from other AI systems of sufficient quality like Stable Diffusion and Midjourney (MJ v5 and above). As users typically… See the full description on the dataset page: https://huggingface.co/datasets/lingcarzy/synthetic-dataset-1m-dalle3-high-quality-captions.Mono-InternVL-2B-Synthetic-Data
Mono-InternVL-2B Synthetic Data
This dataset is used for training the S1.2 stage of Mono-InternVL-2B, as described in the paper Mono-InternVL-1.5: Towards Cheaper and Faster Monolithic Multimodal Large Language Models.
Project Page: https://internvl.github.io/blog/2024-10-10-Mono-InternVL/
Code: https://github.com/OpenGVLab/Mono-InternVL
Dataset Description
Purpose
This dataset is used for training the S1.2 stage of Mono-InternVL-2B.
Data… See the full description on the dataset page: https://huggingface.co/datasets/OpenGVLab/Mono-InternVL-2B-Synthetic-Data.synthetic-derm-1MSynthetic_RCD_2
Synthetic RCD Dataset 2 (CNAM-CD)
This is a synthetic change detection dataset for training Referring Change Detection models like RCDNet.
Important: Pre-change Images (A/)
This dataset does NOT include the pre-change images (A/) because they come from the original CNAM-CD dataset. You have two options:
Option 1: Use Original Dataset (Recommended)
Download the original CNAM-CD dataset from: https://github.com/Silvestezhou/CNAM-CD
Create a symlink to the A/… See the full description on the dataset page: https://huggingface.co/datasets/yilmazkorkmaz/Synthetic_RCD_2.synthetic_jawi_imagesSyntheticFurGroundtruth
Processed Synthetic Fur Dataset
This dataset is part of the dataset located here https://github.com/google-research-datasets/synthetic-fur, with modifications to keep only the ground truth samples and add correponding description ending in *.txt.
Dataset Details
Contains high-quality synthetic rendered fur in static and moving scenes, with different lightning conditions.
synthetic_dimers
Dyno Synthetic Dimer Dataset
The datasets in this repo were used to train Dyno Psi-0 and Dyno Psi-1. This dataset was curated using predicted protein structures from the AlphaFold Protein Structure Database, subset to entries from AFDB50. Dimers were curated by using domain annotations from the TED domain annotations provided by the CATH database to identify domains within the same monomer structure that share an interface. These were further clustered using Foldseek to produce… See the full description on the dataset page: https://huggingface.co/datasets/dynotx/synthetic_dimers.synthetic-vocal-burstsimproved-synthetic-vocal-burtsimproved_synthetic_vocal_burtslogos_synthetic_2synthetic-derm-10ksynthetic-image-caption-pairsmogen3d-synthetic
MoGen3D Synthetic Dataset
Download and Extract
# Download
huggingface-cli download FrankTudor/mogen3d-synthetic synthetic_v2.tar.gz --repo-type dataset --local-dir .
# Extract
tar -xzvf synthetic_v2.tar.gz
Dataset Structure
1000 samples in synthetic_v2/sample_XXXX/ directories
Each sample contains:
metadata.json - Animation info, bone structure
poses_2d.npz - 2D joint positions per frame
poses_3d.npz - 3D joint positions per frame
frame_XXXX.png - Rendered… See the full description on the dataset page: https://huggingface.co/datasets/FrankTudor/mogen3d-synthetic.improved_synthetic_vocal_burtssynthetic-utterancessynthetic-forehead-creasessynthetic-dataset-v2
