CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Narsil /image_dummy\audion<1K0 likes146k downloads5y agoHugging Face02hf-internal-testing /imagefolder_with_metadataimagen<1K0 likes55k downloads2y agoHugging Face03clip-benchmark /wds_imagenet_sketchimage10K<n<100K1 likes19k downloads4y agoHugging Face04imageomics /TreeOfLife-200M Dataset Card for TreeOfLife-200M If you are looking for the original release TreeOfLife-200M dataset, as used in training BioCLIP 2 and presented the paper, please see Revision a8f38b4. The dataset, as presented here, was used to train BioCLIP 2.5 Huge; it completes the dataset cleaning process and resolves an issue where Observation.org occurrences were not included in the training data. With 233 million images representing 933,798 taxa across the tree of life, TreeOfLife-200M… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/TreeOfLife-200M.imageimage-classification100M<n<1B41 likes17k downloads4mo agoHugging Face05imageomics /fish-vista Dataset Card for Fish-Visual Trait Analysis (Fish-Vista) Note that the '</Use this dataset>' option will only load the CSV files. To download the entire dataset, including all processed images and segmentation annotations, refer to Instructions for downloading dataset and images. See Example Code to Use the Segmentation Dataset Figure 1. A schematic representation of the different tasks in Fish-Vista Dataset. Instructions for downloading dataset… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/fish-vista.imageimage-classification10K<n<100K26 likes16k downloads9mo agoHugging Face06adams-story /imagenet1k-256-wdsThis is imagenet1k in webdataset format. Images are stored as jpg files. Every image has been resized to a maximum side length of 256. That means that if an image in the original dataset was 1000 by 500, the new size will be 256 by 128. Images with a maximum side length of under 256 were not resized. The total size of all dataset files is 57.8 GB, there are 1,281,167 rows in the training split and 50,000 rows in the validation split. imageimage-classification100K<n<1M2 likes15k downloads1y agoHugging Face07axiong /imagenet-r ImageNet-R This repo is made to facilitate the evaluation of various pretraining models. It's constructed from the source file provided by official implementation. Usage from datasets import load_dataset dataset = load_dataset('axiong/imagenet-r') Dataset Summary ImageNet-R(endition) contains art, cartoons, deviantart, graffiti, embroidery, graphics, origami, paintings, patterns, plastic objects, plush objects, sculptures, sketches, tattoos, toys, and video… See the full description on the dataset page: https://huggingface.co/datasets/axiong/imagenet-r.image10K<n<100K2 likes15k downloads2y agoHugging Face08trl-internal-testing /zen-imageimagen<1K0 likes15k downloads7mo agoHugging Face09vaishaal /ImageNetV2image10K<n<100K9 likes14k downloads4y agoHugging Face10hf-internal-testing /dummy_image_text_data Dataset Card for "dummy_image_text_data" More Information needed imagen<1K1 likes12k downloads4y agoHugging Face11clip-benchmark /wds_imagenet-rimage10K<n<100K0 likes10k downloads4y agoHugging Face12gmongaras /Imagenet21KNOTE: I have recaptioned all images here This dataset is the entire 21K ImageNet dataset with about 13 million examples and about 19 thousand classes as strings (for some reason it only had ~19K classes instead of 21K). The images are in PNG format. They can be decoded like in the following example import io from PIL import Image Image.open(io.BytesIO(row["image"])) where row["image"] are the raw image bytes. image10M<n<100M9 likes9.2k downloads2y agoHugging Face13trl-internal-testing /zen-multi-imageimagen<1K1 likes9k downloads3mo agoHugging Face14lioooox /T2I-CoReBench-Images T2I-CoReBench-Images 📖 Overview T2I-CoReBench-Images is the companion image dataset of T2I-CoReBench. It contains images generated using 1,080 challenging prompts, covering both composition and reasoning scenarios undere real-world complexities. This dataset is designed to evaluate how well current Text-to-Image (T2I) models can not only paint (produce visually consistent outputs) but also think (perform reasoning over causal chains, object relations, and logical… See the full description on the dataset page: https://huggingface.co/datasets/lioooox/T2I-CoReBench-Images.imagetext-to-image10K<n<100K5 likes8.8k downloads7mo agoHugging Face15gasstation /gs-images-v3tabular100K<n<1M0 likes8.6k downloads5mo agoHugging Face16clip-benchmark /wds_imagenet-aimage1K<n<10K0 likes8.2k downloads4y agoHugging Face17ambrosfitz /19c_newspapers_images_altotabular100K<n<1M4 likes7.7k downloads3mo agoHugging Face18gasstation /gs-images-v2image100K<n<1M1 likes7.7k downloads8mo agoHugging Face19biglam /british-library-book-images British Library Book Images 1,080,814 images cut out of 49,455 digitised books (65,227 volumes, ~25 million pages) published between c. 1510 and c. 1900, digitised by the British Library in partnership with Microsoft and released by British Library Labs on Flickr Commons as the "1 Million Images from Scanned Books" release. The books cover geography, philosophy, history, poetry and literature, in several languages. The four image types British Library Labs… See the full description on the dataset page: https://huggingface.co/datasets/biglam/british-library-book-images.imageimage-classification1M<n<10M64 likes7.6k downloads1mo agoHugging Face20Qwen /Qwen-Image-Bench Qwen-Image-Bench A creator-centric benchmark for evaluating Text-to-Image models beyond semantic alignment. Links Resource Link 📑 Paper http://arxiv.org/abs/2605.28091 📊 Benchmark Dataset (HuggingFace) https://huggingface.co/datasets/Qwen/Qwen-Image-Bench 📊 Benchmark Dataset (ModelScope) https://www.modelscope.cn/datasets/Qwen/Qwen-Image-Bench 💻 GitHub https://github.com/QwenLM/Qwen-Image-Bench 🧑‍⚖️ Q-Judger Model… See the full description on the dataset page: https://huggingface.co/datasets/Qwen/Qwen-Image-Bench.imageimage-to-text1K<n<10K49 likes7.4k downloads4mo agoHugging Face21nvidia /Nemotron-Image-Training-v3 Nemotron Image Training v3 Versions Date Commit Changes 2026-04-28 HEAD Initial commit. Dataset Description Nemotron Image Training v3 is a collection of image-centric multimodal training data for vision–language models. Similar to Nemotron-VLM-Dataset v2, it was curated as a large-scale, multi-subdataset release where each subset ships a standardized conversation JSONL alongside a dataset card describing sources, licensing, and media layout.… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-Image-Training-v3.textvisual-question-answering1M<n<10M82 likes7.1k downloads5mo agoHugging Face22licyk /image_training_set自用的训练集合集,用于 Stable Diffusion 模型微调。 该仓库仅用于存档,不提供任何技术支持。 imagen<1K2 likes7.1k downloads21d agoHugging Face23UCSC-VLAA /GPT-Image-Edit-1.5M GPT-Image-Edit-1.5M A Million-Scale, GPT-Generated Image Dataset 📃Arxiv | 🌐 Project Page | 💻Github GPT-Image-Edit-1.5M is a comprehensive image editing dataset that is built upon HQ-Edit, UltraEdit, OmniEdit and Complex-Edit, with all output images regenerated with GPT-Image-1. 📣 News [2025.08.20] 🚀 We provide a script for multi-process downloading. See Multi-process Download. [2025.07.27] 🤗 We release GPT-Image-Edit, a state-of-the-art image editing model with… See the full description on the dataset page: https://huggingface.co/datasets/UCSC-VLAA/GPT-Image-Edit-1.5M.imageimage-to-image1M<n<10M90 likes7.1k downloads1y agoHugging Face24clip-benchmark /wds_imagenetv2image10K<n<100K0 likes6.7k downloads4y agoHugging Face25lingamvamshikrishnareddy /ramanv-image-captions-realtext100K<n<1M0 likes6.7k downloads24d agoHugging Face26hbXNov /hle_math_exact_match_no_image_int_answerimagen<1K1 likes6.7k downloads2y agoHugging Face27gasstation /gs-images-v4tabular100K<n<1M2 likes6k downloads35m agoHugging Face28hbXNov /hle_math_exact_match_no_image_int_answer_random128imagen<1K0 likes5.9k downloads2y agoHugging Face29timm /imagenet-22k-wdsgated Dataset Summary This is a copy of the full ImageNet dataset consisting of all of the original 21841 clases. It also contains labels in a separate field for the '12k' subset described at at (https://github.com/rwightman/imagenet-12k, https://huggingface.co/datasets/timm/imagenet-12k-wds) This dataset is from the original fall11 ImageNet release which has been replaced by the winter21 release which removes close to 3000 synsets containing people, a number of these are of an offensive… See the full description on the dataset page: https://huggingface.co/datasets/timm/imagenet-22k-wds.imageimage-classification100K<n<1M14 likes5.9k downloads3y agoHugging Face30taesiri /imagenet_hard_review_data_r2tabular1K<n<10K0 likes5.8k downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.