datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
xhs-travel-photos
XHS Travel Photos
Travel photography from Xiaohongshu (Little Red Book) across 19 destinations in China and Southeast Asia.
Dataset Summary
Metric
Count
Notes
2,882
Images
28,005
Total size
7.3 GB
Keyword folders
19
Destinations
Folder
Notes
bali
190
cambodia_angkor
200
chengdu
20
chiang_mai
213
chongqing
194
guilin
225
hainan_sanya
212
indonesia
215
laos
209
malaysia
198
myanmar
122
philippines
210… See the full description on the dataset page: https://huggingface.co/datasets/Rabornkraken/xhs-travel-photos.getting-started-labeled-photos
Dataset Card for predicted_labels
These photos are used in the FiftyOne getting started webinar. The images have a prediction label where were generated by
self-supervised classification through a OpenClip Model.
https://github.com/thesteve0/fiftyone-getting-started/blob/main/5_generating_labels.py
They were then manually cleaned to produce the ground truth label.
https://github.com/thesteve0/fiftyone-getting-started/blob/main/6_clean_labels.md
They are 300 public domain photos… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/getting-started-labeled-photos.french-lot-department-captioned-photos
Lot Department, France Image Dataset
A collection of high-resolution scenic photographs from the Lot region of France with AI-generated descriptive captions.
Dataset Summary
This dataset contains scenic photographs from three notable locations in France's Lot department: Rocamadour, Autoire, and Padirac. All images were captured using a Sony A6600 camera and are paired with detailed English captions generated by Mistral AI's Pixtral-Large model.
Key Features:… See the full description on the dataset page: https://huggingface.co/datasets/NoeFlandre/french-lot-department-captioned-photos.ne_plant_photos
NE Plant Photos
Plant photographs from the iNaturalist open data
archive, filtered to a New England bounding box (latitude 41 to 48, longitude
-74 to -67), each labelled with the species it records and where it was
observed.
This is the nature portion of
blambert/ne_plant_classes -
photos a vision language model judged to show a plant in the field, rather than
a person, a microscope image, or something manmade - so photos of people and
indoor shots are largely gone, but the… See the full description on the dataset page: https://huggingface.co/datasets/blambert/ne_plant_photos.Selfie_and_Official_ID_Photo_Dataset12,000+ people, 150,000+ images. Selfie with ID dataset for KYC verification, face identification and biometric training. Selfies paired with 2 official ID photos (passport, ID card, driver's license, residence permit). 10-15 photos per person with balanced demographics across ethnicity (Caucasian, Black, Asian, Latin American), gender and age (18-65).
Contact us and share your feedback - recieve additional samples for free! 😊
Key Highlights:
12,000+ real individuals… See the full description on the dataset page: https://huggingface.co/datasets/AxonData/Selfie_and_Official_ID_Photo_Dataset.caucasian-people-kyc-photo-dataset
Know Your Customer Dataset, Face Detection and Re-identification
The dataset is created on the basis of Selfies and ID Dataset
80,000+ photos including 10,600+ document photos from 5,300 people from 28 countries.
The dataset includes 2 photos of a person from his documents and 13 selfies. All people presented in the dataset are caucasian. The dataset contains a variety of images capturing individuals from diverse backgrounds and age groups.
Photo documents contains only… See the full description on the dataset page: https://huggingface.co/datasets/UniqueData/caucasian-people-kyc-photo-dataset.vine_virus_photo_dataset
Vine Virus Photo Dataset
A dataset for disease classification of vine plants. The dataset contains 3,866 images across 4 classes: Leafroll 3, No Virus, Other Red, Red Blotch.Images per class:
Leafroll 3: 872
No Virus: 1,825
Other Red: 272
Red Blotch: 897
This dataset is indexed on https://project-agml.github.io/ as part of the AgML python library.
albi-captioned-photos
Albi, France Image Dataset
A collection of high-resolution scenic photographs from Albi, France with AI-generated descriptive captions.
Dataset Summary
This dataset contains scenic photographs from Albi, France, including the city center, the Toulouse Lautrec museum, and the Sainte-Cécile Cathedral. All images were captured using a Sony A6600 camera and are paired with detailed English captions generated by Mistral AI's Pixtral-Large model.
Key Features:
High-resolution… See the full description on the dataset page: https://huggingface.co/datasets/NoeFlandre/albi-captioned-photos.Coral_snake_mimicry_FUNED_photos
Dataset Card for Coral Snake Mimicry specimens from FUNED
This dataset comprises images of snake specimens belonging to species within the coral snake mimicry complex from the Fundacao Ezequiel Dias (FUNED) in Belo Horizonte, MG, Brasil. Images of museum specimens were taken with a Nikon camera and a standardized color palette, but white balance has not been corrected on these images. This dataset is a small representation of a larger dataset collected by Andressa Viol for her PhD… See the full description on the dataset page: https://huggingface.co/datasets/philodryas/Coral_snake_mimicry_FUNED_photos.hispanic-kyc-photo-dataset
Know Your Customer Dataset, Face Detection and Re-identification, Hispanic People
The dataset is created on the basis of Selfies and ID Dataset
10,800+ photos including 1,400+ document photos from 720+ people from 19 countries.
The dataset includes 2 photos of a person from his/her documents and 13 selfies. All people presented in the dataset are hispanic. The dataset contains a variety of images capturing individuals from diverse backgrounds and age groups.
Photo… See the full description on the dataset page: https://huggingface.co/datasets/UniqueData/hispanic-kyc-photo-dataset.Israel-Photos
Israel Photos Dataset
A collection of 369 photographs captured across Israel between 2024 and 2025, with LLM-generated captions and location annotations. The images are sourced from the photographer's Pexels gallery.
About This Collection
This dataset was deliberately curated to provide a diverse visual representation of Israel, encompassing:
Varied locations: From the historic streets of Jerusalem's Old City to Tel Aviv's urban landscape, desert vistas in the Negev, and… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/Israel-Photos.photo-test
Anime vs. Live Action Film Image Classification
OmarK211/photo-test
A binary image classification dataset designed to distinguish between Anime (Label 0) and Live Action (Label 1) film frames/imagery. Images are prepared as standardized square RGB files with multi-pass synthetic training variants.
Source and task
The images were manually imported from the web into my local computer. Then manually uploaded to Google Colab
Preparation source: 24-679 Image Data… See the full description on the dataset page: https://huggingface.co/datasets/OmarK211/photo-test.ARIA_CVI_AAC_Photoset
CVI AAC Photoset
A free, photo-realistic image library for children with Cortical Visual Impairment (CVI) who use Augmentative and Alternative Communication (AAC) boards.
Cortical Visual Impairment is the leading cause of childhood visual impairment in the developed world. Existing AAC vocabulary libraries (ARASAAC, Mulberry, OpenSymbols, SymbolStix, Boardmaker, etc.) are pictogram-style — line drawings on white or colored backgrounds — which the child's cortical visual processing… See the full description on the dataset page: https://huggingface.co/datasets/SkyWhal3/ARIA_CVI_AAC_Photoset.asian-kyc-photo-dataset
Know Your Customer Dataset, Face Detection and Re-identification, Asian People
The dataset is created on the basis of Selfies and ID Dataset
9,900+ photos including 1,300+ document photos from 660 people from 27 countries.
The dataset includes 2 photos of a person from his documents and 13 selfies. All people presented in the dataset are South Asian, East Asian or Middle Asian. The dataset contains a variety of images capturing individuals from diverse backgrounds and… See the full description on the dataset page: https://huggingface.co/datasets/UniqueData/asian-kyc-photo-dataset.QIT-CEMC-FLUTE-PHOTOS
Roles
Roles: canon repo — annot is the source label, kept machine-parseable as the gold for verification and reward parsing; there is no filled reasoning column and this repo is not itself a training view. Derived repos each state their own regime on their own card.
QIT-CEMC 四刃铣刀刃口显微照片(三分类,带 VBmax)
544 张铣刀刃口显微照片,三分类标签(正常 / 磨损 / 毛边卷刃),每张附带上游用显微镜
测量软件量到的 VBmax(后刀面磨损带宽度,毫米)。
照片来自 QIT-CEMC 数据集(Qilu Institute of Technology, Coated End Milling Cutter,… See the full description on the dataset page: https://huggingface.co/datasets/AI4Manufacturing/QIT-CEMC-FLUTE-PHOTOS.vintage-photography-450k-high-quality-captionsThis is a 450k image datastet focused on photography from the 20th century, and their analog aspect. Many of the images are in high resolution. This dataset currently has 20k images captioned with InternVL2 26B, and is a work in progress (I plan to caption the entire dataset and also have short captions for all of the images, compute is an issue for now).
Photo_Dataset
24-679 (Fall 2026): Stairs and Non-Stair Images
ArinRoths/Photo_Dataset
Photos of stairs and non-stairs scenes, prepared as square RGB images with multiple separately
generated training variants. The goal of this dataset is to classify whether or not stairs are
present in an image.
Source and task
The original dataset contains 32 images that I collected and organized into stairs and
non_stairs folders. There are 16 original stairs images and 16 original non-stairs… See the full description on the dataset page: https://huggingface.co/datasets/ArinRoths/Photo_Dataset.MV-VDB-photos-small
MV-VDB-photos-small
Media Vault - Vector Database Photos (Small)
A curated collection of 11,000 images from various computer vision datasets, designed for testing internal mechanisms in the Media Vault Vector Database system. This is the first small-scale dataset (targeting 10K samples, with NSFW split totaling 11K) for validation and testing purposes.
Dataset Structure
The dataset contains two splits:
sfw: All non-NSFW images (~10,000 images)
x_nsfw: Only NSFW images… See the full description on the dataset page: https://huggingface.co/datasets/SamoXXX/MV-VDB-photos-small.vintage-photography-450k-high-quality-captionsThis is a 450k image datastet focused on photography from the 20th century, and their analog aspect. Many of the images are in high resolution. This dataset currently has 20k images captioned with InternVL2 26B, and is a work in progress (I plan to caption the entire dataset and also have short captions for all of the images, compute is an issue for now).
