datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Stable-diffusiondiffusiondbDiffusionDB is the first large-scale text-to-image prompt dataset. It contains 2
million images generated by Stable Diffusion using prompts and hyperparameters
specified by real users. The unprecedented scale and diversity of this
human-actuated dataset provide exciting research opportunities in understanding
the interplay between prompts and generative models, detecting deepfakes, and
designing human-AI interaction tools to help users more easily use these models.prof_report__wavymulder-Analog-Diffusion__multi__24
Dataset Card for "prof_report__wavymulder-Analog-Diffusion__multi__24"
More Information needed
Stable-Diffusion-Prompts
Stable Diffusion Dataset
This is a set of about 80,000 prompts filtered and extracted from the image finder for Stable Diffusion: "Lexica.art". It was a little difficult to extract the data, since the search engine still doesn't have a public API without being protected by cloudflare.
If you want to test the model with a demo, you can go to: "spaces/Gustavosta/MagicPrompt-Stable-Diffusion".
If you want to see the model, go to: "Gustavosta/MagicPrompt-Stable-Diffusion".
rdt-ft-data
Dataset Card
This is the fine-tuning dataset used in the paper RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.
Source
Project Page: https://rdt-robotics.github.io/rdt-robotics/
Paper: https://arxiv.org/pdf/2410.07864
Code: https://github.com/thu-ml/RoboticsDiffusionTransformer
Model: https://huggingface.co/robotics-diffusion-transformer/rdt-1b
Uses
Download all archive files and use the following command to extract:
cat rdt_data.tar.gz.* |… See the full description on the dataset page: https://huggingface.co/datasets/robotics-diffusion-transformer/rdt-ft-data.DiffusionRS-Diffusionprof_images_blip__22h-vintedois-diffusion-v0-1
Dataset Card for "prof_images_blip__22h-vintedois-diffusion-v0-1"
More Information needed
prof_report__22h-vintedois-diffusion-v0-1__multi__24
Dataset Card for "prof_report__22h-vintedois-diffusion-v0-1__multi__24"
More Information needed
stable-diffusion-v1-5-glazed
Dataset Card for Stable Diffusion v1.5 Glazed Samples
Dataset Description
Dataset Summary
This dataset contains image samples originally generated by runwayml/stable-diffusion-v1-5
and subsequently processed by Glaze tool.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information… See the full description on the dataset page: https://huggingface.co/datasets/hanamizuki-ai/stable-diffusion-v1-5-glazed.modelsprof_images_blip__stabilityai-stable-diffusion-2
Dataset Card for "prof_images_blip__stabilityai-stable-diffusion-2"
More Information needed
DiffusionGS
[ICCV 2025] DiffusionGS: Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction
Data Description
These are the demo results of our ICCV 2025 paper.
HuggingFace Model Link
We also release our models in HuggingFace:
https://huggingface.co/CaiYuanhao/DiffusionGS
Here are some video generation results demo:
· (a) Object-level Generation… See the full description on the dataset page: https://huggingface.co/datasets/CaiYuanhao/DiffusionGS.stable-diffusion-backup
Stable Diffusion web UI
A browser interface based on Gradio library for Stable Diffusion.
Features
Detailed feature showcase with images:
Original txt2img and img2img modes
One click install and run script (but you still must install python and git)
Outpainting
Inpainting
Color Sketch
Prompt Matrix
Stable Diffusion Upscale
Attention, specify parts of text that the model should pay more attention to
a man in a ((tuxedo)) - will pay more attention to tuxedo
a man in a… See the full description on the dataset page: https://huggingface.co/datasets/artdwn/stable-diffusion-backup.stable-diffusion-webui-reForge
Stable Diffusion WebUI Forge/reForge
Stable Diffusion WebUI Forge/reForge is a platform on top of Stable Diffusion WebUI (based on Gradio) to make development easier, optimize resource management, speed up inference, and study experimental features.
The name "Forge" is inspired from "Minecraft Forge". This project is aimed at becoming SD WebUI's Forge.
Forge2/reForge2
You can read more on… See the full description on the dataset page: https://huggingface.co/datasets/WhiteAiZ/stable-diffusion-webui-reForge.text2dataset
Text2Dataset
the dataset generate from ChatGPT's output and text2light for training the LoRA using in DiffusionLight
Code for training the LoRA can be found at DiffusionLight-LoRA-Trainer
OAS95-aligned-cleaneddiffusion-pretrain-set-ft1
diffusion-pretrain-set-ft1
A multi-source image-caption pretraining dataset assembled from ten upstream
sources via a uniform ingest pipeline. Designed for a full pretrain or finetune
pipeline meant to curate for any major diffusion model preliminary, with the sole
intent to create a more powerful baseline preliminary train and a baseline
for synthesizing images to train the next generation of the VLM model.
This is a lot like the snake eating it's own tail, so it must be… See the full description on the dataset page: https://huggingface.co/datasets/AbstractPhil/diffusion-pretrain-set-ft1.stable-diffusion-webui
Stable Diffusion web UI
A browser interface based on Gradio library for Stable Diffusion.
Features
Detailed feature showcase with images:
Original txt2img and img2img modes
One click install and run script (but you still must install python and git)
Outpainting
Inpainting
Color Sketch
Prompt Matrix
Stable Diffusion Upscale
Attention, specify parts of text that the model should pay more attention to
a man in a ((tuxedo)) - will pay more attention to tuxedo
a man in a… See the full description on the dataset page: https://huggingface.co/datasets/PennyJX/stable-diffusion-webui.Diffusion4D-Animated-Mocap-7440
Diffusion4D Animated Mocap 7440
Rolling export of 7,440 technically and semantically selected animated assets.
Each asset has 12 yaw views (0..330 degrees, step 30), 24 sampled frames,
BVH motion, DINOv2 image embeddings, and one PNG preview. Batch archives are
uploaded only after local stage validation.
Stable_Diffusion_3_RecaptionThis dataset is the one specified in the stable diffusion 3 paper which is composed of the ImageNet dataset and the CC12M dataset.
I used the ImageNet 2012 train/val data and captioned it as specified in the paper: "a photo of a 〈class name〉" (note all ids are 999,999,999)
CC12M is a dataset with 12 million images created in 2021. Unfortunately the downloader provided by Google has many broken links and the download takes forever.
However, some people in the community publicized the dataset.… See the full description on the dataset page: https://huggingface.co/datasets/gmongaras/Stable_Diffusion_3_Recaption.prof_report__CompVis-stable-diffusion-v1-4__multi__24
Dataset Card for "prof_report__CompVis-stable-diffusion-v1-4__multi__24"
More Information needed
Diffusion-book-cn
《从零开始学扩散模型》
术语表
词汇
翻译
Corruption Process
退化过程
Pipeline
管线
Timestep
时间步
Scheduler
调度器
Gradient Accumulation
梯度累加
Fine-Tuning
微调
Guidance
引导
目录
第一部分 基础知识
第一章 扩散模型的原理、发展和应用
1.1 扩散模型的原理
1.2 扩散模型的发展
1.3 扩散模型的应用
第二章 HuggingFace介绍与环境准备
2.1 HuggingFace Space
2.2 Transformer 与 diffusers 库
2.3 环境准备… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFace-CN-community/Diffusion-book-cn.Diffusion4D-Animated-Raw
Diffusion4D Animated Assets
This dataset provides animated 3D assets referenced by Diffusion4D and
Objaverse-XL in a directly browsable format. The default split contains 67,988
rows. Each row includes metadata, a preview image, and a short preview video so
that assets can be inspected in the Hugging Face Data Studio without first
downloading the original 3D file.
The repository also mirrors available raw assets and keeps their original
source links and hashes. The current… See the full description on the dataset page: https://huggingface.co/datasets/DenisKochetov/Diffusion4D-Animated-Raw.blip3o-256diffusion_721_clipped_4K_10B_sampledFrom67Bgray_scott_reaction_diffusionThis Dataset is part of The Well Collection.
How To Load from HuggingFace Hub
Be sure to have the_well installed (pip install the_well)
Use the WellDataModule to retrieve data as follows:
from the_well.data import WellDataModule
# The following line may take a couple of minutes to instantiate the datamodule
datamodule = WellDataModule(
"hf://datasets/polymathic-ai/",
"gray_scott_reaction_diffusion",
)
train_dataloader = datamodule.train_dataloader()
for batch in… See the full description on the dataset page: https://huggingface.co/datasets/polymathic-ai/gray_scott_reaction_diffusion.GenRef-wds
GenRef-1M
We provide 1M high-quality triplets of the form (flawed image, high-quality image, reflection) collected across
multiple domains using our scalable pipeline from [1]. We used this dataset to train our reflection tuning model.
To know the details of the dataset creation pipeline, please refer to Section 3.2 of [1].
Project Page: https://diffusion-cot.github.io/reflection2perfection
Dataset loading
We provide the dataset in the webdataset format for fast… See the full description on the dataset page: https://huggingface.co/datasets/diffusion-cot/GenRef-wds.diffusion_policy_robocasa_activations_latest_chkpt
Diffusion Policy — RoboCasa Activations (latest checkpoint)
Per-step, per-episode activation traces collected from a DiffusionTransformerHybridImagePolicy (the diffusion_policy library) rolled out on RoboCasa benchmark tasks. Captured with collect_activations_robocasa.py at the latest training checkpoint.
These traces are the input expected by the conceptor / SAE steering pipelines under diffusion_policy/experiments/robocasa_steering/ and diffusion_policy/experiments/sae/ — see… See the full description on the dataset page: https://huggingface.co/datasets/brandonyang/diffusion_policy_robocasa_activations_latest_chkpt.Stable-Diffusion
