CoolFace
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01diffusion-bench /blip3o-256image1K<n<10K1 likes1.2k downloads6mo agoHugging Face02diffusion-cot /GenRef-wds GenRef-1M We provide 1M high-quality triplets of the form (flawed image, high-quality image, reflection) collected across multiple domains using our scalable pipeline from [1]. We used this dataset to train our reflection tuning model. To know the details of the dataset creation pipeline, please refer to Section 3.2 of [1]. Project Page: https://diffusion-cot.github.io/reflection2perfection Dataset loading We provide the dataset in the webdataset format for fast… See the full description on the dataset page: https://huggingface.co/datasets/diffusion-cot/GenRef-wds.imagetext-to-image1M<n<10M15 likes1.1k downloads1y agoHugging Face03diffusion-cot /GenRef-CoT GenRef-CoT We provide 227K high-quality CoT reflections which were used to train our Qwen-based reflection generation model in ReflectionFlow [1]. To know the details of the dataset creation pipeline, please refer to Section 3.2 of [1]. Dataset loading We provide the dataset in the webdataset format for fast dataloading and streaming. We recommend downloading the repository locally for faster I/O: from huggingface_hub import snapshot_download local_dir =… See the full description on the dataset page: https://huggingface.co/datasets/diffusion-cot/GenRef-CoT.image100K<n<1M3 likes464 downloads1y agoHugging Face04dhyun22 /video-diffusion-perceptionimage10M<n<100M0 likes327 downloads6mo agoHugging Face05robotics-diffusion-transformer /BimanualUR5eExample Dataset Summary This dataset provides shards in the WebDataset format for fine-tuning RDT-2 or other policy models on bimanual manipulation. Each sample packs: a binocular RGB image (left + right wrist cameras concatenated horizontally) a relative action chunk (continuous control, 0.8s, 30Hz) a discrete action token sequence (e.g., from an Residual VQ action tokenizer) a metadata JSON with an instruction key sub_task_instruction_key to index corresponding instruction from… See the full description on the dataset page: https://huggingface.co/datasets/robotics-diffusion-transformer/BimanualUR5eExample.imagerobotics100K<n<1M3 likes85 downloads8mo agoHugging Face06hmu013 /DiffusionDB-300k-processedimage100K<n<1M0 likes78 downloads8mo agoHugging Face07fabiotosi92 /Diffusion4RobustDepth Diffusion4RobustDepth This repository contains the generated dataset and trained network weights used in the paper "Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditions" (ECCV 2024). Dataset Structure The dataset is organized into three main categories: driving/: Contains autonomous driving datasets with challenging images. ToM/: Contains the Transparent and Mirrored (ToM) objects dataset. weights/: Contains the weights of models trained in… See the full description on the dataset page: https://huggingface.co/datasets/fabiotosi92/Diffusion4RobustDepth.imagedepth-estimation100K<n<1M2 likes65 downloads2y agoHugging Face08hmu013 /DiffusionDB-300kimage100K<n<1M0 likes26 downloads9mo agoHugging Face09omkarthawakar /stable_diffusion_dataimage10K<n<100K0 likes16 downloads1y agoHugging Face10dhyun22 /video-diffusion-perception-feasibilityimage10K<n<100K1 likes6 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.