datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
gs-images-v319c_newspapers_images_altobritish-library-book-images
British Library Book Images
1,080,814 images cut out of 49,455 digitised books (65,227 volumes, ~25 million pages) published
between c. 1510 and c. 1900, digitised by the British Library in partnership
with Microsoft and released by British Library Labs
on Flickr Commons as the "1 Million Images from Scanned Books" release. The books cover geography,
philosophy, history, poetry and literature, in several languages.
The four image types
British Library Labs… See the full description on the dataset page: https://huggingface.co/datasets/biglam/british-library-book-images.prof_images_blip__22h-vintedois-diffusion-v0-1
Dataset Card for "prof_images_blip__22h-vintedois-diffusion-v0-1"
More Information needed
prof_images_blip__stabilityai-stable-diffusion-2
Dataset Card for "prof_images_blip__stabilityai-stable-diffusion-2"
More Information needed
british-library-book-images
British Library Book Images
1,080,814 images cut out of 49,455 digitised books (65,227 volumes, ~25 million pages) published
between c. 1510 and c. 1900, digitised by the British Library in partnership
with Microsoft and released by British Library Labs
on Flickr Commons as the "1 Million Images from Scanned Books" release. The books cover geography,
philosophy, history, poetry and literature, in several languages.
The four image types
British Library Labs… See the full description on the dataset page: https://huggingface.co/datasets/Faizaniqbal/british-library-book-images.dummy-base64-imagesprof_images_blip__SG161222-Realistic_Vision_V1.4
Dataset Card for "prof_images_blip__SG161222-Realistic_Vision_V1.4"
More Information needed
open-imagesgenerated-imageschartqa_without_images
Dataset Card for "chartqa_without_images"
If you wanna load the dataset, you can run the following code:
from datasets import load_dataset
data = load_dataset('ahmed-masry/chartqa_without_images')
The dataset has the following structure:
DatasetDict({
train: Dataset({
features: ['imgname', 'query', 'label', 'type'],
num_rows: 28299
})
val: Dataset({
features: ['imgname', 'query', 'label', 'type'],
num_rows: 1920
})
test:… See the full description on the dataset page: https://huggingface.co/datasets/ahmed-masry/chartqa_without_images.open-images-v7-subsetFineVision_imagesCOCO2014-Images
Dataset Card for "COCO2014-Images"
More Information needed
bbc_images_alltime
RealTimeData Monthly Collection - BBC News Images
This datasets contains all news articles head images from BBC News that were created every months from 2017 to current.
To access articles in a specific month, simple run the following:
ds = datasets.load_dataset('RealTimeData/bbc_images_alltime', '2020-02')
This will give you all BBC news head images that were created in 2020-02.
Want to crawl the data by your own?
Please head to LatestEval for the crawler… See the full description on the dataset page: https://huggingface.co/datasets/RealTimeData/bbc_images_alltime.fashion-product-images-small
Dataset Card for "fashion-product-images-small"
More Information needed
Data was obtained from here
AToMiC-Images-v0.2
Dataset Card for "AToMiC-All-Images_wi-pixels"
Languages
The dataset contains 108 languages in Wikipedia.
Data Instances
Each instance is an image, its representation in bytes, and its associated captions.
Intended Usage
Image collection for Text-to-Image retrieval
Image--Caption Retrieval/Generation/Translation
Licensing Information
CC BY-SA 4.0 international license
Citation Information
TBA
Acknowledgement
Thanks… See the full description on the dataset page: https://huggingface.co/datasets/TREC-AToMiC/AToMiC-Images-v0.2.prof_images_blip__andite-anything-v4.0
Dataset Card for "prof_images_blip__andite-anything-v4.0"
More Information needed
AI-Generated-vs-Real-Images-Datasets
Dataset Card for "AI-Generated-vs-Real-Images-Datasets"
More Information needed
telugu-synthetic-line-imagesprof_images_blip__andite-pastel-mix
Dataset Card for "prof_images_blip__andite-pastel-mix"
More Information needed
pixmo_images
PixMo Images
The raw images backing the PixMo
datasets used to train Molmo, packaged
as Parquet shards with embedded image bytes so they can be browsed in the dataset viewer
and loaded directly with datasets.
The PixMo annotation datasets (allenai/pixmo-*) ship image_urls rather than image
bytes. This repository is a content cache of those images, keyed by the SHA-256 of the
source URL.
Contents
1,073,189 images across 525 Parquet shards (data/train-*.parquet)… See the full description on the dataset page: https://huggingface.co/datasets/UWGZQ/pixmo_images.popsign-images
PopSign Images Dataset
This dataset contains frame sequences extracted from PopSign ASL (American Sign Language) video clips, organized for sign language recognition tasks.
Dataset Description
The PopSign dataset consists of short video clips of isolated ASL signs. This version provides pre-extracted image frames from each video clip, suitable for training image-based or video-based models for sign language recognition.
Subsets
The dataset contains two subsets:… See the full description on the dataset page: https://huggingface.co/datasets/sign/popsign-images.human-coherence-preferences-images
Rapidata Image Generation Coherence Dataset
This dataset was collected in ~4 Days using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation.
Explore our latest model rankings on our website.
If you get value from this dataset and would like to see more in the future, please consider liking it.
Overview
One of the largest human annotated coherence datasets for text-to-image models, this release contains over 1,200,000 human… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/human-coherence-preferences-images.dhivehi-vrd-images
Dhivehi Single-Line Text-Image Dataset
A collection of synthetic Dhivehi text images for training and evaluating text-image / vision models etc. Each image contains a single line of Dhivehi text with various visual styles and augmentations (Check the config field for more info on the row).
Batch Statistics
Batch
Total Images
Train
Validation
Test
vrd-batch-1
474169
379335
47417
47417
vrd-batch-2
474493
379594
47449
47450
vrd-batch-3
475564
380451
47556… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/dhivehi-vrd-images.pixmo-cap-images
PixMo-Cap
Big thanks to Ai2 for releasing the original PixMo-Cap dataset. To preserve the images and simplify usage of the dataset, we are releasing this version, which includes downloaded images.
PixMo-Cap is a dataset of very long (roughly 200 words on average), detailed captions.
It can be used to pre-train and fine-tune vision-language models.
PixMo-Cap was created by recording annotators speaking about an image for 60-90 seconds and then using the Claude large language model… See the full description on the dataset page: https://huggingface.co/datasets/anthracite-org/pixmo-cap-images.prof_images_blip__CompVis-stable-diffusion-v1-4
Dataset Card for "prof_images_blip__CompVis-stable-diffusion-v1-4"
More Information needed
laion_improved_aesthetics_6.5plus_with_imageshuman-alignment-preferences-images
Rapidata Image Generation Alignment Dataset
This dataset was collected in ~4 Days using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation.
Explore our latest model rankings on our website.
If you get value from this dataset and would like to see more in the future, please consider liking it.
Overview
One of the largest human annotated alignment datasets for text-to-image models, this release contains over 1,200,000 human… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/human-alignment-preferences-images.pixmo-cap-images
