datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
imagenet1k-256-wdsThis is imagenet1k in webdataset format. Images are stored as jpg files. Every image has been resized to a maximum side length of 256. That means that if an image in the original dataset was 1000 by 500, the new size will be 256 by 128. Images with a maximum side length of under 256 were not resized.
The total size of all dataset files is 57.8 GB, there are 1,281,167 rows in the training split and 50,000 rows in the validation split.
stable-diffusion-v1-5-glazed
Dataset Card for Stable Diffusion v1.5 Glazed Samples
Dataset Description
Dataset Summary
This dataset contains image samples originally generated by runwayml/stable-diffusion-v1-5
and subsequently processed by Glaze tool.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information… See the full description on the dataset page: https://huggingface.co/datasets/hanamizuki-ai/stable-diffusion-v1-5-glazed.StreetView360AtoZStreetView 360X is a dataset containing 6342 360 degree equirectangular street view images randomly sampled and downloaded from Google Street View. It is published as part of the paper "StreetView360X: A Location-Conditioned Latent Diffusion Model for Generating Equirectangular 360 Degree Street Views" (Princeton COS Senior Independent Work by Everett Shen). Images are labelled with their capture coordinates and panorama IDs. Scripts for extending the dataset (i.e. fetching additional images)… See the full description on the dataset page: https://huggingface.co/datasets/everettshen/StreetView360AtoZ.jev-stage2-image-beans-pilot
Beans: one natural question per image
Open the corrected preview.
natural_v4 is the recommended and default preview: 100 original images, 100 rows, one three-way condition-class Choice question per image. All targets come directly from the source labels column (34 angular leaf spot, 33 bean rust, 33 healthy). Original image bytes and source annotations are unchanged.
Example question: “Which source-defined condition class describes the bean leaf?” Options: angular_leaf_spot… See the full description on the dataset page: https://huggingface.co/datasets/FaroukMoc2/jev-stage2-image-beans-pilot.world-streetview-500k
🌍 World StreetView 500k
World StreetView 500k is a large-scale computer vision dataset for visual geolocation estimation, spatial representation learning, and geographic scene understanding.
It pairs ~2 million street-level images from ~500,000 unique locations worldwide with geographic coordinates, country labels, capture dates, and elevation data. Each training location is captured from 4 compass headings (0°, 90°, 180°, 270°) — ideal for training GeoGuessr-style geolocation… See the full description on the dataset page: https://huggingface.co/datasets/josefbednar/world-streetview-500k.human-style-preferences-images
Rapidata Image Generation Preference Dataset
This dataset was collected in ~4 Days using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation.
Explore our latest model rankings on our website.
If you get value from this dataset and would like to see more in the future, please consider liking it.
Overview
One of the largest human preference datasets for text-to-image models, this release contains over 1,200,000 human preference… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/human-style-preferences-images.STHELAR_40x
STHELAR dataset (40x)
STHELAR (Spatial Transcriptomics and H&E histology for Large-scale Annotation Resource) is a multi-tissue dataset designed for developing models capable of predicting cell types directly from histological Hematoxylin & Eosin (H&E) whole slide images. It integrates high-resolution spatial transcriptomics data with histology, to provide detailed segmentation masks and cell-type annotations.
Available dataset versions
STHELAR_40x — 587,555 image… See the full description on the dataset page: https://huggingface.co/datasets/FelicieGS/STHELAR_40x.STRI-Samples
Dataset Card for Smithsonian Tropical Research Institute (STRI) Samples
Dataset Summary
Dorsal images of butterfly wings collected by Owen McMillan and members of his lab at the Smithsonian Tropical Research Institute.
Full dataset will be 24,119 RGB images: Dorsal and Ventral images of separated wings. This sample contains 207 dorsal butterfly images used as part of the training data for Imageomics/butterfly_detection_yolo.
Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/STRI-Samples.STHELAR_20x
STHELAR dataset (20x)
STHELAR (Spatial Transcriptomics and H&E histology for Large-scale Annotation Resource) is a multi-tissue dataset designed for developing models capable of predicting cell types directly from histological Hematoxylin & Eosin (H&E) whole slide images. It integrates high-resolution spatial transcriptomics data with histology, to provide detailed segmentation masks and cell-type annotations.
Available dataset versions
STHELAR_40x — 587,555 image… See the full description on the dataset page: https://huggingface.co/datasets/FelicieGS/STHELAR_20x.random_streetview_images_pano_v0.0.2
Dataset Card for panoramic street view images (v.0.0.2)
Dataset Summary
The random streetview images dataset are labeled, panoramic images scraped from randomstreetview.com. Each image shows a location
accessible by Google Streetview that has been roughly combined to provide ~360 degree view of a single location. The dataset was designed with the intent to geolocate an image purely based on its visual content.
Supported Tasks and Leaderboards
None as of now!… See the full description on the dataset page: https://huggingface.co/datasets/stochastic/random_streetview_images_pano_v0.0.2.prl_crimson_maiden_style
prl_crimson_maiden_style Dataset
This dataset is used to train my lora Crimson Maiden from Civitai.
painting-style-classification
Dataset Labels
['Realism', 'Art_Nouveau_Modern', 'Analytical_Cubism', 'Cubism', 'Expressionism', 'Action_painting', 'Synthetic_Cubism', 'Symbolism', 'Ukiyo_e', 'Naive_Art_Primitivism', 'Post_Impressionism', 'Impressionism', 'Fauvism', 'Rococo', 'Minimalism', 'Mannerism_Late_Renaissance', 'Color_Field_Painting', 'High_Renaissance', 'Romanticism', 'Pop_Art', 'Contemporary_Realism', 'Baroque', 'New_Realism', 'Pointillism', 'Northern_Renaissance', 'Early_Renaissance'… See the full description on the dataset page: https://huggingface.co/datasets/keremberke/painting-style-classification.tactile-mnist-touch-starstruck-syn-single-t32-320x240Documentation is available at https://github.com/TimSchneider42/tactile-mnist/blob/main/doc/datasets.md#touch-datasets.
Canadian-streetview-cities
Canadian Street View Cities Dataset
Overview
A street-view image dataset created to train and evaluate models for city-level image classification across major Canadian cities. Each entry includes an image and its corresponding city label.
Purpose
The dataset is intended for building models that recognize the Canadian city in which a street-view scene was captured.
Data Source
All images were collected from Mapillary, using geographic bounding boxes… See the full description on the dataset page: https://huggingface.co/datasets/SABR22/Canadian-streetview-cities.stem-diagrams
STEM Diagrams
30,325 technical diagrams (block diagrams, schematics, flowcharts, architectures)
extracted from arXiv papers across six engineering fields, each with a source
attribution and a quality score. Built by an LLM-curated pipeline and used to show
that a small frozen-feature classifier can replace the paid LLM labeling gate.
Paper: Distilling an LLM Diagram-Curation Pipeline into Local Classifiers (Adnan Abbasi, Thothica, 2026)
Code:… See the full description on the dataset page: https://huggingface.co/datasets/aeyxen/stem-diagrams.typhoon-intensity-classification
Typhoon - Image Classification Dataset
This dataset comes from PTIT AI Challenge and is organized for a multi-class image classification task focusing on tropical cyclone (typhoon) intensity estimation.
Dataset Structure
The directory structure is organized as follows:
train/
├── images/
│ ├── image1.jpg
│ └── ...
└── annotations.csv (only present in the train folder)
The public_test and private_test sets are used to evaluate and score the… See the full description on the dataset page: https://huggingface.co/datasets/star092304/typhoon-intensity-classification.Strawberry-MM-Straw5
Dataset Card for Strawberry Disease Multimodal Dataset
Dataset Description
This is a multimodal dataset for strawberry disease detection, which contains strawberry image data, corresponding environmental parameters (air temperature, air humidity, soil moisture) and strawberry variety information. It can be used to study the correlation between environmental factors and strawberry disease occurrence, as well as multimodal fusion disease detection algorithms.… See the full description on the dataset page: https://huggingface.co/datasets/Qin2006/Strawberry-MM-Straw5.NSFW-MultiDomain-Classification
NSFW_MultiDomain
The NSFW_MultiDomain dataset is a curated image classification dataset focused on multi-domain adult content recognition. It consists of 5 distinct categories aimed at facilitating the development of robust NSFW (Not Safe For Work) image classification models. This dataset enables training and benchmarking of models that can distinguish between subtle variations in explicit and non-explicit content across artistic, animated, and real-world imagery.… See the full description on the dataset page: https://huggingface.co/datasets/strangerguardhf/NSFW-MultiDomain-Classification.plism-dataset-tiles-st
Stain-Transferred PLISM Dataset
Dataset Overview
The Stain-Transferred PLISM dataset is a synthetic variant of the PLISM dataset tiles provided by Filiot et al. (2025), which is based on the original PLISM-wsi dataset by Ochi et al. (2024).
This dataset isolates global color-level variations (staining profiles) from localized morphological and scanner-specific hardware artifacts. It is specifically designed to evaluate and robustify computational pathology… See the full description on the dataset page: https://huggingface.co/datasets/klotz11/plism-dataset-tiles-st.pokemon_card_image_for_authenticity_classification
Pokemon Card Image for Authenticity Classification
This dataset contains front/back images of Pokemon cards for authenticity experiments.
Dataset structure
Images/: all image files (.jpeg)
Images/metadata.jsonl: metadata used by Hugging Face imagefolder
labels.csv: flat label file with the same rows as metadata
Columns
image: image object loaded from file
id: image filename (unique id)
side: card side (0 = front, 1 = back)
labels: authenticity label (1 =… See the full description on the dataset page: https://huggingface.co/datasets/stevelohwc/pokemon_card_image_for_authenticity_classification.strawberry_growth_detection
Strawberry Growth Detection
A dataset for detection of strawberry growth stages. The dataset contains 1,477 images with 3,997 bounding box annotations across 7 categories. The dataset also contains ground truth data
related to the size of the strawberries from tagged leaves, as well as a decimal based growth stage.
This dataset is indexed on https://project-agml.github.io/ as part of the AgML python library.
Citation
@article{yang2024predicting… See the full description on the dataset page: https://huggingface.co/datasets/Project-AgML/strawberry_growth_detection.StreetView-Image-Dataset-10K
Urban Streetscape Dataset for Vision Language Models
A curated subset of 10,000 street view images with 25 essential features optimized for training vision language models on urban environment analysis tasks.
Dataset Description
This dataset contains street view imagery paired with comprehensive annotations covering infrastructure characteristics, visual perception metrics, environmental context, and semantic segmentation data.
This comprehensive dataset represents a… See the full description on the dataset page: https://huggingface.co/datasets/Sadhana-24/StreetView-Image-Dataset-10K.stock-charts
Stock Charts
This dataset is a collection of a sample of images from tweets that I scraped using my Discord bot that keeps track of financial influencers on Twitter.
The data consists of images that were part of tweets that mentioned a stock.
This dataset can be used for a wide variety of tasks, such as image classification or feature extraction.
FinTwit Charts Collection
This dataset is part of a larger collection of datasets, scraped from Twitter and labeled by a… See the full description on the dataset page: https://huggingface.co/datasets/StephanAkkerman/stock-charts.streetview-global
StreetView Global
A globally-sampled street-view image dataset with rich scene annotations and
visual question-answer pairs. All photographs are sourced from
Mapillary, the open street-level imagery
platform, via its public image API.
Each example pairs a street-level photograph with geographic metadata
(latitude, longitude, compass, capture time, region), a free-form scene
description, structured scene classification (setting, weather, time of day,
road type, infrastructure), and… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/streetview-global.StairvsNonStair_Dataset
24-679 (Fall 2026): Stairs and Non-Stair Images
ArinRoths/StairvsNonStair_Dataset
Photos of stairs and non-stairs scenes, prepared as square RGB images with multiple separately
generated training variants. The goal of this dataset is to classify whether or not stairs are
present in an image.
Source and task
The original dataset contains 32 images that I collected and organized into stairs and
non_stairs folders. There are 16 original stairs images and 16 original… See the full description on the dataset page: https://huggingface.co/datasets/ArinRoths/StairvsNonStair_Dataset.sd2-staged-foreign-objects
SD2 staged-laboratory foreign-object frames (heydonto)
97 frames · 163 frame-level annotation rows · 8 staged laboratory takes · CC BY 4.0
This dataset discloses and carries the SD2 own-footage portion of the training data of the ORena SAVE FOCUS challenge entry's FRAME-track component: 163 of that component's 58,086 pooled training rows. The other sources of that corpus are not part of this dataset.
What is in it
frames/ — 97 JPEG frames (filename = first 16 hex… See the full description on the dataset page: https://huggingface.co/datasets/HeyDonto/sd2-staged-foreign-objects.StreetView360AtoZStreetView 360X is a dataset containing 6342 360 degree equirectangular street view images randomly sampled and downloaded from Google Street View. It is published as part of the paper "StreetView360X: A Location-Conditioned Latent Diffusion Model for Generating Equirectangular 360 Degree Street Views" (Princeton COS Senior Independent Work by Everett Shen). Images are labelled with their capture coordinates and panorama IDs. Scripts for extending the dataset (i.e. fetching additional images)… See the full description on the dataset page: https://huggingface.co/datasets/vinod18vin/StreetView360AtoZ.hmdb51-pick-run-stand
HMDB51 — Pick / Run / Stand (frames procesados)
Subconjunto procesado del dataset HMDB51 para un caso de uso de
clasificación de productividad de empleados en almacén mediante visión por
computadora: distinguir entre trabajador activo (recogiendo / corriendo)
e inactivo/pausado (parado).
Clases
Clase HMDB51
Etiqueta de negocio
pick
Activo (Pick Up)
run
Activo (Running)
stand
Inactivo/Pausado (Standing)
Estadísticas
492 videos… See the full description on the dataset page: https://huggingface.co/datasets/treborDev/hmdb51-pick-run-stand.StatMapCorpus
StatMapCorpus v1
23,549 English-language statistical maps identified in MapPool, annotated for cartographic
method by vision-language models.
This repository is a mirror. The citable source of record is the deposit in Dane
Badawcze UW (University of Warsaw, ICM): https://doi.org/10.58132/FSJGSP, version 1.0.
The nine files in deposit/ are byte-identical to that release; deposit/schema.json
carries their sha256 sums so you can verify this yourself. Cite the DOI, not this URL.… See the full description on the dataset page: https://huggingface.co/datasets/msolarz/StatMapCorpus.crypto-charts
Crypto Charts
This dataset is a collection of a sample of images from tweets that I scraped using my Discord bot that keeps track of financial influencers on Twitter.
The data consists mainly of images that are cryptocurrency charts.
This dataset can be used for a wide variety of tasks, such as image classification or feature extraction.
FinTwit Charts Collection
This dataset is part of a larger collection of datasets, scraped from Twitter and labeled by a human (me).… See the full description on the dataset page: https://huggingface.co/datasets/StephanAkkerman/crypto-charts.
