CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01gasstation /gs-images-v3tabular100K<n<1M0 likes9k downloads5mo agoHugging Face02ambrosfitz /19c_newspapers_images_altotabular100K<n<1M4 likes8.1k downloads3mo agoHugging Face03biglam /british-library-book-images British Library Book Images 1,080,814 images cut out of 49,455 digitised books (65,227 volumes, ~25 million pages) published between c. 1510 and c. 1900, digitised by the British Library in partnership with Microsoft and released by British Library Labs on Flickr Commons as the "1 Million Images from Scanned Books" release. The books cover geography, philosophy, history, poetry and literature, in several languages. The four image types British Library Labs… See the full description on the dataset page: https://huggingface.co/datasets/biglam/british-library-book-images.imageimage-classification1M<n<10M64 likes6.2k downloads1mo agoHugging Face04sasha /prof_images_blip__22h-vintedois-diffusion-v0-1 Dataset Card for "prof_images_blip__22h-vintedois-diffusion-v0-1" More Information needed image10K<n<100K0 likes3.4k downloads3y agoHugging Face05sasha /prof_images_blip__stabilityai-stable-diffusion-2 Dataset Card for "prof_images_blip__stabilityai-stable-diffusion-2" More Information needed image10K<n<100K0 likes2.7k downloads3y agoHugging Face06Faizaniqbal /british-library-book-images British Library Book Images 1,080,814 images cut out of 49,455 digitised books (65,227 volumes, ~25 million pages) published between c. 1510 and c. 1900, digitised by the British Library in partnership with Microsoft and released by British Library Labs on Flickr Commons as the "1 Million Images from Scanned Books" release. The books cover geography, philosophy, history, poetry and literature, in several languages. The four image types British Library Labs… See the full description on the dataset page: https://huggingface.co/datasets/Faizaniqbal/british-library-book-images.imageimage-classification1M<n<10M0 likes2.6k downloads1mo agoHugging Face07hf-internal-testing /dummy-base64-imagestextn<1K0 likes2.5k downloads2y agoHugging Face08sasha /prof_images_blip__SG161222-Realistic_Vision_V1.4 Dataset Card for "prof_images_blip__SG161222-Realistic_Vision_V1.4" More Information needed image10K<n<100K0 likes2.3k downloads3y agoHugging Face09dalle-mini /open-imagesimage1M<n<10M27 likes2k downloads5y agoHugging Face10gasstation /generated-imagesimage100K<n<1M0 likes1.9k downloads10mo agoHugging Face11ahmed-masry /chartqa_without_images Dataset Card for "chartqa_without_images" If you wanna load the dataset, you can run the following code: from datasets import load_dataset data = load_dataset('ahmed-masry/chartqa_without_images') The dataset has the following structure: DatasetDict({ train: Dataset({ features: ['imgname', 'query', 'label', 'type'], num_rows: 28299 }) val: Dataset({ features: ['imgname', 'query', 'label', 'type'], num_rows: 1920 }) test:… See the full description on the dataset page: https://huggingface.co/datasets/ahmed-masry/chartqa_without_images.text10K<n<100K1 likes1.8k downloads3y agoHugging Face12bitmind /open-images-v7-subsetimage1M<n<10M0 likes1.8k downloads2y agoHugging Face13lightonai /FineVision_imagesimage1M<n<10M1 likes1.8k downloads1y agoHugging Face14justram /COCO2014-Images Dataset Card for "COCO2014-Images" More Information needed image100K<n<1M2 likes1.8k downloads3y agoHugging Face15RealTimeData /bbc_images_alltime RealTimeData Monthly Collection - BBC News Images This datasets contains all news articles head images from BBC News that were created every months from 2017 to current. To access articles in a specific month, simple run the following: ds = datasets.load_dataset('RealTimeData/bbc_images_alltime', '2020-02') This will give you all BBC news head images that were created in 2020-02. Want to crawl the data by your own? Please head to LatestEval for the crawler… See the full description on the dataset page: https://huggingface.co/datasets/RealTimeData/bbc_images_alltime.image100K<n<1M2 likes1.7k downloads1y agoHugging Face16ashraq /fashion-product-images-small Dataset Card for "fashion-product-images-small" More Information needed Data was obtained from here image10K<n<100K45 likes1.6k downloads4y agoHugging Face17TREC-AToMiC /AToMiC-Images-v0.2gated Dataset Card for "AToMiC-All-Images_wi-pixels" Languages The dataset contains 108 languages in Wikipedia. Data Instances Each instance is an image, its representation in bytes, and its associated captions. Intended Usage Image collection for Text-to-Image retrieval Image--Caption Retrieval/Generation/Translation Licensing Information CC BY-SA 4.0 international license Citation Information TBA Acknowledgement Thanks… See the full description on the dataset page: https://huggingface.co/datasets/TREC-AToMiC/AToMiC-Images-v0.2.image10M<n<100M4 likes1.5k downloads4y agoHugging Face18sasha /prof_images_blip__andite-anything-v4.0 Dataset Card for "prof_images_blip__andite-anything-v4.0" More Information needed image1K<n<10K0 likes1.3k downloads3y agoHugging Face19Hemg /AI-Generated-vs-Real-Images-Datasets Dataset Card for "AI-Generated-vs-Real-Images-Datasets" More Information needed image100K<n<1M25 likes1.2k downloads3y agoHugging Face20harsha-desaraju /telugu-synthetic-line-imagesimage1M<n<10M0 likes1.1k downloads2mo agoHugging Face21sasha /prof_images_blip__andite-pastel-mix Dataset Card for "prof_images_blip__andite-pastel-mix" More Information needed image10K<n<100K0 likes1k downloads3y agoHugging Face22UWGZQ /pixmo_images PixMo Images The raw images backing the PixMo datasets used to train Molmo, packaged as Parquet shards with embedded image bytes so they can be browsed in the dataset viewer and loaded directly with datasets. The PixMo annotation datasets (allenai/pixmo-*) ship image_urls rather than image bytes. This repository is a content cache of those images, keyed by the SHA-256 of the source URL. Contents 1,073,189 images across 525 Parquet shards (data/train-*.parquet)… See the full description on the dataset page: https://huggingface.co/datasets/UWGZQ/pixmo_images.imageimage-to-text1M<n<10M1 likes1k downloads3mo agoHugging Face23sign /popsign-images PopSign Images Dataset This dataset contains frame sequences extracted from PopSign ASL (American Sign Language) video clips, organized for sign language recognition tasks. Dataset Description The PopSign dataset consists of short video clips of isolated ASL signs. This version provides pre-extracted image frames from each video clip, suitable for training image-based or video-based models for sign language recognition. Subsets The dataset contains two subsets:… See the full description on the dataset page: https://huggingface.co/datasets/sign/popsign-images.imagevisual-question-answering100K<n<1M0 likes1k downloads8mo agoHugging Face24Rapidata /human-coherence-preferences-images Rapidata Image Generation Coherence Dataset This dataset was collected in ~4 Days using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation. Explore our latest model rankings on our website. If you get value from this dataset and would like to see more in the future, please consider liking it. Overview One of the largest human annotated coherence datasets for text-to-image models, this release contains over 1,200,000 human… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/human-coherence-preferences-images.imagetext-to-image10K<n<100K14 likes1k downloads2y agoHugging Face25alakxender /dhivehi-vrd-images Dhivehi Single-Line Text-Image Dataset A collection of synthetic Dhivehi text images for training and evaluating text-image / vision models etc. Each image contains a single line of Dhivehi text with various visual styles and augmentations (Check the config field for more info on the row). Batch Statistics Batch Total Images Train Validation Test vrd-batch-1 474169 379335 47417 47417 vrd-batch-2 474493 379594 47449 47450 vrd-batch-3 475564 380451 47556… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-vrd-images.imagevisual-question-answering1M<n<10M0 likes987 downloads1y agoHugging Face26anthracite-org /pixmo-cap-images PixMo-Cap Big thanks to Ai2 for releasing the original PixMo-Cap dataset. To preserve the images and simplify usage of the dataset, we are releasing this version, which includes downloaded images. PixMo-Cap is a dataset of very long (roughly 200 words on average), detailed captions. It can be used to pre-train and fine-tune vision-language models. PixMo-Cap was created by recording annotators speaking about an image for 60-90 seconds and then using the Claude large language model… See the full description on the dataset page: https://huggingface.co/datasets/anthracite-org/pixmo-cap-images.imageimage-to-text100K<n<1M13 likes923 downloads2y agoHugging Face27sasha /prof_images_blip__CompVis-stable-diffusion-v1-4 Dataset Card for "prof_images_blip__CompVis-stable-diffusion-v1-4" More Information needed image10K<n<100K0 likes915 downloads3y agoHugging Face28bhargavsdesai /laion_improved_aesthetics_6.5plus_with_imagestext100K<n<1M23 likes846 downloads4y agoHugging Face29Rapidata /human-alignment-preferences-images Rapidata Image Generation Alignment Dataset This dataset was collected in ~4 Days using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation. Explore our latest model rankings on our website. If you get value from this dataset and would like to see more in the future, please consider liking it. Overview One of the largest human annotated alignment datasets for text-to-image models, this release contains over 1,200,000 human… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/human-alignment-preferences-images.imagetext-to-image10K<n<100K17 likes757 downloads2y agoHugging Face30dnth /pixmo-cap-imagesimage100K<n<1M1 likes754 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.