CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Gustavosta /Stable-Diffusion-Prompts Stable Diffusion Dataset This is a set of about 80,000 prompts filtered and extracted from the image finder for Stable Diffusion: "Lexica.art". It was a little difficult to extract the data, since the search engine still doesn't have a public API without being protected by cloudflare. If you want to test the model with a demo, you can go to: "spaces/Gustavosta/MagicPrompt-Stable-Diffusion". If you want to see the model, go to: "Gustavosta/MagicPrompt-Stable-Diffusion". text10K<n<100K527 likes5.3k downloads4y agoHugging Face02hanamizuki-ai /stable-diffusion-v1-5-glazed Dataset Card for Stable Diffusion v1.5 Glazed Samples Dataset Description Dataset Summary This dataset contains image samples originally generated by runwayml/stable-diffusion-v1-5 and subsequently processed by Glaze tool. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information… See the full description on the dataset page: https://huggingface.co/datasets/hanamizuki-ai/stable-diffusion-v1-5-glazed.imageimage-classification100K<n<1M3 likes2.9k downloads3y agoHugging Face03bayes-group-diffusion /OAS95-aligned-cleanedtext100M<n<1B0 likes1.8k downloads10mo agoHugging Face04AbstractPhil /diffusion-pretrain-set-ft1 diffusion-pretrain-set-ft1 A multi-source image-caption pretraining dataset assembled from ten upstream sources via a uniform ingest pipeline. Designed for a full pretrain or finetune pipeline meant to curate for any major diffusion model preliminary, with the sole intent to create a more powerful baseline preliminary train and a baseline for synthesizing images to train the next generation of the VLM model. This is a lot like the snake eating it's own tail, so it must be… See the full description on the dataset page: https://huggingface.co/datasets/AbstractPhil/diffusion-pretrain-set-ft1.image1M<n<10M2 likes1.8k downloads3mo agoHugging Face05PennyJX /stable-diffusion-webui Stable Diffusion web UI A browser interface based on Gradio library for Stable Diffusion. Features Detailed feature showcase with images: Original txt2img and img2img modes One click install and run script (but you still must install python and git) Outpainting Inpainting Color Sketch Prompt Matrix Stable Diffusion Upscale Attention, specify parts of text that the model should pay more attention to a man in a ((tuxedo)) - will pay more attention to tuxedo a man in a… See the full description on the dataset page: https://huggingface.co/datasets/PennyJX/stable-diffusion-webui.imagen<1K0 likes1.6k downloads3y agoHugging Face06gmongaras /Stable_Diffusion_3_RecaptionThis dataset is the one specified in the stable diffusion 3 paper which is composed of the ImageNet dataset and the CC12M dataset. I used the ImageNet 2012 train/val data and captioned it as specified in the paper: "a photo of a 〈class name〉" (note all ids are 999,999,999) CC12M is a dataset with 12 million images created in 2021. Unfortunately the downloader provided by Google has many broken links and the download takes forever. However, some people in the community publicized the dataset.… See the full description on the dataset page: https://huggingface.co/datasets/gmongaras/Stable_Diffusion_3_Recaption.image10M<n<100M5 likes1.3k downloads2y agoHugging Face07DenisKochetov /Diffusion4D-Animated-Raw Diffusion4D Animated Assets This dataset provides animated 3D assets referenced by Diffusion4D and Objaverse-XL in a directly browsable format. The default split contains 67,988 rows. Each row includes metadata, a preview image, and a short preview video so that assets can be inspected in the Hugging Face Data Studio without first downloading the original 3D file. The repository also mirrors available raw assets and keeps their original source links and hashes. The current… See the full description on the dataset page: https://huggingface.co/datasets/DenisKochetov/Diffusion4D-Animated-Raw.3d10K<n<100K4 likes1.3k downloads3mo agoHugging Face08diffusion-bench /blip3o-256image1K<n<10K1 likes1.2k downloads6mo agoHugging Face09diffusion-cot /GenRef-wds GenRef-1M We provide 1M high-quality triplets of the form (flawed image, high-quality image, reflection) collected across multiple domains using our scalable pipeline from [1]. We used this dataset to train our reflection tuning model. To know the details of the dataset creation pipeline, please refer to Section 3.2 of [1]. Project Page: https://diffusion-cot.github.io/reflection2perfection Dataset loading We provide the dataset in the webdataset format for fast… See the full description on the dataset page: https://huggingface.co/datasets/diffusion-cot/GenRef-wds.imagetext-to-image1M<n<10M15 likes1.1k downloads1y agoHugging Face10rbeauchamp /diffusion_db_dedupe_from50k_train Dataset Card for "diffusion_db_dedupe_from50k_train" More Information needed image10K<n<100K1 likes952 downloads3y agoHugging Face11AbstractPhil /diffusion-pretrain-set-ft1-1024 diffusion-pretrain-set-ft1-1024 1024px (2x) upscale of AbstractPhil/diffusion-pretrain-set-ft1. WARNING MUCH OF THIS DATA WAS MODEL UPSCALED USING RAPID UPSCALERS. THIS IS NOT CONSISTENTLY HIGH FIDELITY NOR IS IT EVEN CLOSE TO FAIR FIDELITY AT TIMES. PLEASE use this ONLY for pretraining, new concepts, and simple design purposes ONLY. HEAVILY PRUNE FOR FINETUNING. Thank you, good luck my friends. Details Model: realesr-general-x4v3 (SRVGG Compact… See the full description on the dataset page: https://huggingface.co/datasets/AbstractPhil/diffusion-pretrain-set-ft1-1024.image1M<n<10M0 likes882 downloads3mo agoHugging Face12DiffusionArcade /Pongimage10K<n<100K0 likes760 downloads1y agoHugging Face13giakoupg /diffusionprint_dataset DiffusionPrint Patch Dataset A dataset of 64x64 image patches for contrastive learning of diffusion-based inpainting forensics. Each patch comes from either a real image or an AI-inpainted region generated by one of three diffusion models. Columns Column Type Description image bytes (PNG) Lossless 64x64 RGB patch master_index int64 Row index in the original memmap archive patch_path string Original relative path of the patch category string real… See the full description on the dataset page: https://huggingface.co/datasets/giakoupg/diffusionprint_dataset.textimage-classification1M<n<10M0 likes614 downloads2mo agoHugging Face14fffffchopin /DiffusionDream_DatasetThis is the dataset of the diffusion dream dataset. The dataset contains the following columns: info: A string describing the action taken in the frame keyword: A string describing the keyword of the action action: A string describing the action taken in the frame current_frame: The current frame of the video previous_frame_1: The frame before the current frame previous_frame_2: The frame before the previous frame previous_frame_3: The frame before the previous frame previous_frame_4: The… See the full description on the dataset page: https://huggingface.co/datasets/fffffchopin/DiffusionDream_Dataset.image100K<n<1M1 likes594 downloads1y agoHugging Face15gzzyyxy /layout_diffusion_hypersimThis repository contains the data for SceneCraft: Layout-Guided 3D Scene Generation. Project page: https://orangesodahub.github.io/SceneCraft Code: https://github.com/OrangeSodahub/SceneCraft imagetext-to-3d10K<n<100K1 likes550 downloads1y agoHugging Face16jtatman /stable-diffusion-prompts-stats-full-uncensoredimage100K<n<1M151 likes506 downloads2y agoHugging Face17LGirrbach /person-centric-images-stable-diffusion-v1-1image100K<n<1M0 likes492 downloads1y agoHugging Face18diffusion-cot /GenRef-CoT GenRef-CoT We provide 227K high-quality CoT reflections which were used to train our Qwen-based reflection generation model in ReflectionFlow [1]. To know the details of the dataset creation pipeline, please refer to Section 3.2 of [1]. Dataset loading We provide the dataset in the webdataset format for fast dataloading and streaming. We recommend downloading the repository locally for faster I/O: from huggingface_hub import snapshot_download local_dir =… See the full description on the dataset page: https://huggingface.co/datasets/diffusion-cot/GenRef-CoT.image100K<n<1M3 likes464 downloads1y agoHugging Face19svjack /diffusiondb_2m_random_50k Dataset Card for "diffusiondb_2m_random_50k" More Information needed image10K<n<100K0 likes453 downloads4y agoHugging Face20rdoshi21 /1m2r-real-diffusion2This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "kinova-arx", "total_episodes": 72, "total_frames": 32350, "total_tasks": 1, "total_videos": 0, "total_chunks": 1, "chunks_size": 1000, "fps": 10, "splits": { "train": "0:72" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rdoshi21/1m2r-real-diffusion2.imagerobotics10K<n<100K0 likes441 downloads8mo agoHugging Face21yuwan0 /lexica-stable-diffusion-v1-5 Stable Diffusion Dataset This is a set of about 80,000 Image-Prompt pairs generated by stable-diffusion-v1-5. The Prompts come from dataset Stable-Diffusion-Prompts which filtered and extracted from the image finder for Stable Diffusion: "Lexica.art". image10K<n<100K4 likes344 downloads3y agoHugging Face22dhyun22 /video-diffusion-perceptionimage10M<n<100M0 likes327 downloads6mo agoHugging Face23teticio /audio-diffusion-1024Over 20,000 256x256 mel spectrograms of 5 second samples of music from my Spotify liked playlist. The code to convert from audio to spectrogram and vice versa can be found in https://github.com/teticio/audio-diffusion along with scripts to train and run inference using De-noising Diffusion Probabilistic Models. x_res = 1024 y_res = 1024 sample_rate = 44100 n_fft = 2048 hop_length = 512 imageimage-to-image10K<n<100K0 likes299 downloads4y agoHugging Face24ririye /Benchmark-Images-for-Stable-Diffusion-Biastext1K<n<10K0 likes297 downloads2y agoHugging Face25cyttic /diffusionpen-hebrew-handwriting DiffusionPen Hebrew Handwriting A large synthetic dataset of Hebrew handwritten text lines with ground-truth transcriptions, for training and evaluating handwritten text recognition (HTR / OCR) models. Every image is a single line of right-to-left Hebrew handwriting synthesized by DiffusionPen — a style-conditioned latent-diffusion handwriting generator — in one of 491 distinct writer styles, and quality-filtered by an independent OCR pass. 149,952 line images, 491 writer… See the full description on the dataset page: https://huggingface.co/datasets/cyttic/diffusionpen-hebrew-handwriting.imageimage-to-text100K<n<1M1 likes282 downloads3mo agoHugging Face26neuralworm /stable-diffusion-discord-promptsstable-diffusion-discord-prompts All messages from dreambot from all dream-[1-50] channels in stable-diffusion discord source: https://github.com/bartman081523/stable-diffusion-discord-prompts text1M<n<10M29 likes278 downloads4y agoHugging Face27AdrianPrados /DiffusionPolicyMinJerkimagen<1K1 likes270 downloads12d agoHugging Face28ANWERFATEHY /high-quality_art-mix_images_for_diffusion_training_1ai_made photorealistic image100K<n<1M0 likes266 downloads4d agoHugging Face29punwaiw /DiffusionJockey Dataset Card for "DiffusionJockey" More Information needed image1K<n<10K0 likes265 downloads3y agoHugging Face30thefcraft /civitai-stable-diffusion-337k How to Use from datasets import load_dataset dataset = load_dataset("thefcraft/civitai-stable-diffusion-337k") print(dataset['train'][0]) download images download zip files from images dir https://huggingface.co/datasets/thefcraft/civitai-stable-diffusion-337k/tree/main/images it contains some images with id from zipfile import ZipFile with ZipFile("filename.zip", 'r') as zObject: zObject.extractall() Dataset Summary GitHub URL:-… See the full description on the dataset page: https://huggingface.co/datasets/thefcraft/civitai-stable-diffusion-337k.image100K<n<1M43 likes261 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.