datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dummy_image_text_data
Dataset Card for "dummy_image_text_data"
More Information needed
open-image-preferences-v1
Open Image Preferences
Prompt: Anime-style concept art of a Mayan Quetzalcoatl biomutant, dystopian world, vibrant colors, 4K.
Image 1
Image 2
Prompt: 8-bit pixel art of a blue knight, green car, and glacier landscape in Norway, fantasy style, colorful and detailed.
Image 1… See the full description on the dataset page: https://huggingface.co/datasets/data-is-better-together/open-image-preferences-v1.Defactify_Image_Dataset
Defactify_Image_Dataset
This dataset is associated with the paper A Comprehensive Dataset for Human vs. AI Generated Image Detection.
📝 Dataset Description
Dataset Summary
The Defactify_Image_Dataset (A Comprehensive Dataset for Human vs. AI Generated Image Detection) is a high-quality collection of 96,000 images and associated metadata designed to benchmark models for detecting and identifying the source of artificially generated content. Built using the MS… See the full description on the dataset page: https://huggingface.co/datasets/Rajarshi-Roy-research/Defactify_Image_Dataset.dummy_image_class_data
Dataset Card for "dummy_image_class_data"
More Information needed
CIFAKE-image-datasetcat-image-datasetLabels (in this order):
sks cat sitting on a chair in front of a box of chocolates
sks cat playing on the Steam Deck
sks cat wearing a necklace while sitting in a box on a sofa
a box with four donuts in front of sks cat
sks cat wearing a pink veil with flowers on it and a dagger made out of yellow cardboard
sks cat between two pillows, with one pillow showing a polar bear and the other a fox
sks cat with an espresso reading the newspaper
a close up of a hand petting sks cat on the head
sks cat… See the full description on the dataset page: https://huggingface.co/datasets/peft-internal-testing/cat-image-dataset.pagoda-text-and-image-dataset
Dataset Card for "pagoda-text-and-image-dataset"
More Information needed
yt_full_image_dataset
Dataset Card for "yt_full_image_dataset"
More Information needed
image-matching-test-datasetfashion-image-datasettext-2-image-human-preferences-2m
Text-to-image human preferences: 2M votes across 30 models
This dataset contains the complete voting record behind the
Datapoint Image Bench
leaderboard: 2,161,160 validated pairwise votes — exactly 10 for each of
216,116 image pairs. The votes compare 30 text-to-image models in a complete
round-robin on 500 prompts, judged by annotators from over 200 countries.
Every vote includes the annotator's trust score at the time the vote was
cast.
Built on the Datapoint annotation… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-image-human-preferences-2m.open-image-preferences-v1-binarized
Open Image Preferences
Prompt: Anime-style concept art of a Mayan Quetzalcoatl biomutant, dystopian world, vibrant colors, 4K.
Image 1
Image 2
Prompt: 8-bit pixel art of a blue knight, green car, and glacier landscape in Norway, fantasy style, colorful and detailed.
Image 1… See the full description on the dataset page: https://huggingface.co/datasets/data-is-better-together/open-image-preferences-v1-binarized.yt_main_image_dataset
Dataset Card for "yt_main_image_dataset"
More Information needed
waste-dataset-image
♻️ Waste Classification Image Dataset
Dataset Summary
This dataset contains standardized, multi-class waste images categorized across 9 target recycling categories. It was compiled, curated, and manually sanitized to eliminate noise, corrupted files, and class overlap in order to train high-performance convolutional neural networks and transfer learning backbones (such as EfficientNet-B5).
The dataset contains a total of 28,840 labeled images, split into dedicated… See the full description on the dataset page: https://huggingface.co/datasets/4w4kt/waste-dataset-image.seatizen_atlas_image_dataset
Seatizen Atlas Image Dataset
Dataset Card
Dataset Name: Seatizen Atlas Image DatasetTask: Multi-label image classificationDomain: Marine BiodiversityLicense: cc0-1.0Size: 14,492 annotated images
Description
The Seatizen Atlas Image Dataset is a large-scale collection of annotated underwater images designed for training and evaluating artificial intelligence models in marine biodiversity research. It is specifically tailored for multi-label image… See the full description on the dataset page: https://huggingface.co/datasets/lombardata/seatizen_atlas_image_dataset.5190-image-dataset
Dataset
This dataset contains images taken along walkways between 33rd and Walnut to 34th and Spruce St in Philadelphia, PA, USA. Each image is tagged with latitude and longitude labels.
Each sample consists of:
An image
A latitude coordinate
A longitude coordinate
Intended Use
The dataset is intended for the UPenn CIS 4190/5190 Img2GPS project that trains a model to take in an image and output latitude and longitude coordinates.
my-image-caption-datasetimage-text-dataset-subset-300k-captions_onlyimage-matching-datasetgovdocs1-image
BEE-spoke-data/govdocs1-image
This contains .jpg files from govdocs1. Light deduplication was applied (i.e. jdupes on all files) which removed ~500 duplicate images.
DatasetDict({
train: Dataset({
features: ['image'],
num_rows: 108895
})
})
source
Source info/page: https://digitalcorpora.org/corpora/file-corpora/files/
@inproceedings{garfinkel2009bringing,
title={Bringing Science to Digital Forensics with Standardized Forensic Corpora}… See the full description on the dataset page: https://huggingface.co/datasets/BEE-spoke-data/govdocs1-image.pagoda-text-and-image-dataset-small
Dataset Card for "pagoda-text-and-image-dataset-small"
More Information needed
brain-tumor-image-dataset-semantic-segmentation
Dataset Card for "brain-tumor-image-dataset-semantic-segmentation"
Dataset Description
The Brain Tumor Image Dataset (BTID) for Semantic Segmentation contains MRI images and annotations aimed at training and evaluating segmentation models. This dataset was sourced from Kaggle and includes detailed segmentation masks indicating the presence and boundaries of brain tumors.
This dataset can be used for developing and benchmarking algorithms for medical image segmentation… See the full description on the dataset page: https://huggingface.co/datasets/dwb2023/brain-tumor-image-dataset-semantic-segmentation.celeb-df-image-datasetshopping-queries-image-dataset
Shopping Queries Image Dataset (SQID 🦑): An Image-Enriched ESCI Dataset for Exploring Multimodal Learning in Product Search
Introduction
The Shopping Queries Image Dataset (SQID) is a dataset that includes image information for over 190,000 products. This dataset is an augmented version of the Amazon Shopping Queries Dataset, which includes a large number of product search queries from real Amazon users, along with a list of up to 40 potentially relevant results and… See the full description on the dataset page: https://huggingface.co/datasets/crossingminds/shopping-queries-image-dataset.low-alt-satellite-image-dataset-5k-sam3-segmented_jsonmy_image_captioning_datasetwebsite_screenshots_image_dataset
Website Screenshots Image Dataset
This dataset is obtainable here from roboflow..
Dataset Details
Dataset Description
Language(s) (NLP): [English]
License: [MIT]
Dataset Sources
Source: [https://universe.roboflow.com/roboflow-gw7yv/website-screenshots/dataset/1]
Uses
From the roboflow website:
Annotated screenshots are very useful in Robotic Process Automation. But they can be expensive to label. This dataset would cost over… See the full description on the dataset page: https://huggingface.co/datasets/Zexanima/website_screenshots_image_dataset.image-datasetimage_diff_data
VDiff-Bench
VDiff-Bench is a multiple-choice benchmark for fine-grained visual difference identification. Each example presents two similar images and four candidate descriptions, exactly one of which states a real difference between the images.
Dataset structure
The train split contains 1,756 questions with the following fields:
id: stable example identifier.
image_1, image_2: the paired images.
choices: an object containing answer choices A, B, C, and D.… See the full description on the dataset page: https://huggingface.co/datasets/elaine1wan/image_diff_data.low-alt-satellite-image-dataset-5k-sam3-segmented
