datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
fish-vista
Dataset Card for Fish-Visual Trait Analysis (Fish-Vista)
Note that the '</Use this dataset>' option will only load the CSV files. To download the entire dataset, including all processed images and segmentation annotations, refer to Instructions for downloading dataset and images.
See Example Code to Use the Segmentation Dataset
Figure 1. A schematic representation of the different tasks in Fish-Vista Dataset.
Instructions for downloading dataset… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/fish-vista.image-as-an-imu-finetuning
Image as an IMU: Real-world Finetuning Dataset
Official real-world finetuning dataset from Image as an IMU: Estimating Camera Motion from a Single Motion-Blurred Image (ICCV 2025 Oral).
[arXiv] [Webpage] [GitHub]
PIXL, University of Oxford
Jerred Chen, Ronald Clark
Dataset Details
This dataset consists of 32 sequences of real-world motion-blurred videos in various indoor scenes, captured using the iPhone 13 camera.
dataset_train_real-world.csv and… See the full description on the dataset page: https://huggingface.co/datasets/jerredchen00/image-as-an-imu-finetuning.VLM4Bio
Dataset Card for VLM4Bio
Instructions for downloading the dataset
Install Git LFS
Git clone the VLM4Bio repository to download all metadata and associated files
Run the following commands in a terminal:
git clone https://huggingface.co/datasets/imageomics/VLM4Bio
cd VLM4Bio
Downloading and processing bird images
To download the bird images, run the following command:
bash download_bird_images.sh
This should download the bird images inside datasets/Bird/images… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/VLM4Bio.Tunisian-Proverbs-with-Image-Associations-A-Cultural-and-Linguistic-DatasetTunisian Proverbs with Image Associations: A Cultural and Linguistic Dataset
Description
This dataset explores the rich oral tradition of Tunisian proverbs mapped into text format, pairing each with contextual explanations, English translations both word-to-word and it's equivalent Target Language dynamic, Automated prompt and AI-generated visual interpretations.
It bridges linguistic, cultural, and visual modalities making it valuable for tasks in cross-cultural NLP, generative… See the full description on the dataset page: https://huggingface.co/datasets/HabibaAbderrahim/Tunisian-Proverbs-with-Image-Associations-A-Cultural-and-Linguistic-Dataset.questFish2024
Dataset Card for QUEST Fish 2024
Images collected by teachers during a QUEST workshop. In 2024, the images were of fish collected from bodies of water near Princeton University.
Dataset Details
Dataset Structure
/dataset/
<folder>/
<img_id 1>.png
<img_id 2>.png
...
<img_id n>.png
...
<img_id 1>.png
<img_id 2>.png
...
<img_id n>.png
fieldData2024.csv
Data Instances… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/questFish2024.Heliconius-Collection_Cambridge-Butterfly
Dataset Card for Heliconius Collection (Cambridge Butterfly)
Dataset Description
Dataset Summary
Subset of the collection records from Chris Jiggins' research group at the University of Cambridge, collection covers nearly 20 years of field studies.
This subset contains approximately 36,189 RGB images of 11,962 specimens (29,134 images of 10,086 specimens across all Heliconius). Many records have both images and locality data.
Most images were… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/Heliconius-Collection_Cambridge-Butterfly.crowdsourced-sea-images-v2crowdsourced-sea-images
Dataset Card for Dataset Name
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/SEA-AI/crowdsourced-sea-images.FashionGEN_images_datatgk-ai-image-generators-2026
We Tested 10 AI Image Generators on Faces, Text and Ads
Most AI image-generator comparisons reduce the models to a score. We wanted to see the mistakes.
These Guys Know gave ten current models the same three practical briefs in August 2026: a close-up face, exact medical text inside a photographed hospital monitor, and a luxury fragrance advertisement where the person, bottle, label and location needed to look believable together.
We kept the first valid output for every… See the full description on the dataset page: https://huggingface.co/datasets/These-Guys-Know/tgk-ai-image-generators-2026.image-preference-demo
Image dataset for preference aquisition demo
This dataset provides the files used to run the example that we use in this blog post to illustrate how easily
you can set up and run the annotation process to collect a huge preference dataset using Rapidata's API.
The goal is to collect human preferences based on pairwise image matchups.
The dataset contains:
Generated images: A selection of example images generated using Flux.1 and Stable Diffusion. The images are provided in a .zip… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/image-preference-demo.japanese-image-classification-evaluation-dataset
recruit-jp/japanese-image-classification-evaluation-dataset
Overview
Developed by: Recruit Co., Ltd.
Dataset type: Image Classification
Language(s): Japanese
LICENSE: CC-BY-4.0
More details are described in our tech blog post.
日本語CLIP学習済みモデルとその評価用データセットの公開
Dataset Details
This dataset is comprised of four image classification tasks related to concepts and things unique to Japan. Specifically, is consists of the following tasks.
jafood101: Image… See the full description on the dataset page: https://huggingface.co/datasets/recruit-jp/japanese-image-classification-evaluation-dataset.ImageNet-CJ
JPEG Re-encoding Confound Control Dataset
A controlled-experiment dataset that isolates one acknowledged-but-unmeasured confound in
ImageNet-C. Hendrycks & Dietterich (Benchmarking Neural Network Robustness to Common
Corruptions and Perturbations, ICLR 2019, arXiv:1903.12261)
save every corrupted image as a lightly compressed JPEG. The benchmark therefore never measures a
corruption c applied to an image x in isolation — it measures JPEG(c(x)). This dataset lets
you quantify how… See the full description on the dataset page: https://huggingface.co/datasets/atharvadagaonkar/ImageNet-CJ.gpt-image-edit-benchmark-results
GPT-Image-Edit — Benchmark Results
This repository contains evaluation results of GPT-Image-Edit across four standard image-editing benchmarks. All scores were computed using the official evaluation scripts provided by each benchmark.
📊 Benchmarks
Benchmark
Metrics
Folder
GEdit-EN
12 editing categories + Avg
gedit/
Complex-Edit
IF, IP, PQ, Overall
complex_edit/
ImgEdit-Full
10 editing operations + Overall
imgedit/
OmniContext
Contextual edit scores… See the full description on the dataset page: https://huggingface.co/datasets/UCSC-VLAA/gpt-image-edit-benchmark-results.NEON-plant-subplot-pilot
Dataset Card for NEON Plant Presence and Percent Cover Subplot Pilot Images
Pilot dataset for using subplot images from the National Ecological Observatory Network (NEON) to detect plant diversity.
Dataset Details
Dataset Description
Repository: PlotDiversiVision
This dataset contains plant diversity images from two NEON sites: CPER and SCBI.
Each image captures a 1-square-meter subplot labeled with plant species.
NEON curated subplot… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/NEON-plant-subplot-pilot.clinical-narrative-image-integrity-v0.2
Clinical Narrative Image Integrity v0.2
What this is
A small dataset that tests one question:
Can you detect when a clinical narrative-image system is moving toward integrity failure, not just carrying ambiguity?
This repo focuses on narrative-image integrity under clinical reasoning pressure.
It models a system where:
narrative coherence may weaken
image alignment may drift
interpretive distortion may rise
fragmented signal may destabilize representation before overt… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-narrative-image-integrity-v0.2.simple-image-captionsimage-transcreationImages-Datasetchatinterface_with_image_csv
Dataset Card for Dataset Name
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/freddyaboulton/chatinterface_with_image_csv.testJourneyBench_Multi_Image_VQA
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/JourneyBench/JourneyBench_Multi_Image_VQA.bershka-imagesCrop_Disease_Images
Crop Disease Expert Annotations
1,092 crop images annotated by agricultural experts with ground truth diagnoses covering pests, diseases, and nutrient deficiencies across 74 crop types.
Dataset
File: annotations.csv (4 columns)
Column
Description
image_url
URL to the crop image
crop
Crop or plant identified by the expert
diagnosis
Pest, disease, or nutrient deficiency name(s); Healthy if none
details
Expert's observations on visible symptoms… See the full description on the dataset page: https://huggingface.co/datasets/kirutheen/Crop_Disease_Images.Corroboration-Image
Corroboration-Image
tags: multimodal, visual verification, image corroboration
Note: This is an AI-generated dataset so its content may be inaccurate or false
Dataset Description:
The 'Corroboration-Image' dataset is a curated collection of images paired with textual evidence used for training machine learning models in the task of visual verification and image corroboration. The dataset is designed to support multimodal learning approaches, where a model learns to associate textual… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/Corroboration-Image.Fitzwilliam-museum-imagesA CSV file of image urls and meta data for the Fitzwilliam Museum system. The images are licensed under more restrictive terms, the links to URLS
are open via their API and website.
Crop_Disease_Images
Crop Disease Expert Annotations
1,092 crop images annotated by agricultural experts with ground truth diagnoses covering pests, diseases, and nutrient deficiencies across 74 crop types.
Dataset
File: annotations.csv (4 columns)
Column
Description
image_url
URL to the crop image
crop
Crop or plant identified by the expert
diagnosis
Pest, disease, or nutrient deficiency name(s); Healthy if none
details
Expert's observations on visible symptoms… See the full description on the dataset page: https://huggingface.co/datasets/Mithun2009/Crop_Disease_Images.ImageCaptioning_CatalanThe dataset consists of 153,791 images, each accompanied by a description in Catalan. The images have been sourced from two repositories:
"yerevann/coco-karpathy" and "UCSC-VLAA/Recap-COCO-30K." This dataset is ideal for computer vision tasks, as it combines a wide variety of
images with detailed descriptions that can be useful for training machine learning models.
It is freely accessible to everyone, as long as proper credit is given to the original data sources. Thanks
ScienceQA_Image_urlImageClassifyTrain
ImageClassifyTrain
tags: Image Recognition, Classification, Large Scale
Note: This is an AI-generated dataset so its content may be inaccurate or false
Dataset Description:
The 'ImageClassifyTrain' dataset is a collection of images used for training machine learning models in the field of image recognition and classification. Each image is accompanied by a label that categorizes the image into one of several predefined classes. This dataset includes a variety of images sourced from… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/ImageClassifyTrain.
