CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01adams-story /imagenet1k-256-wdsThis is imagenet1k in webdataset format. Images are stored as jpg files. Every image has been resized to a maximum side length of 256. That means that if an image in the original dataset was 1000 by 500, the new size will be 256 by 128. Images with a maximum side length of under 256 were not resized. The total size of all dataset files is 57.8 GB, there are 1,281,167 rows in the training split and 50,000 rows in the validation split. imageimage-classification100K<n<1M2 likes16k downloads1y agoHugging Face02hanamizuki-ai /stable-diffusion-v1-5-glazed Dataset Card for Stable Diffusion v1.5 Glazed Samples Dataset Description Dataset Summary This dataset contains image samples originally generated by runwayml/stable-diffusion-v1-5 and subsequently processed by Glaze tool. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information… See the full description on the dataset page: https://huggingface.co/datasets/hanamizuki-ai/stable-diffusion-v1-5-glazed.imageimage-classification100K<n<1M3 likes2.9k downloads3y agoHugging Face03everettshen /StreetView360AtoZStreetView 360X is a dataset containing 6342 360 degree equirectangular street view images randomly sampled and downloaded from Google Street View. It is published as part of the paper "StreetView360X: A Location-Conditioned Latent Diffusion Model for Generating Equirectangular 360 Degree Street Views" (Princeton COS Senior Independent Work by Everett Shen). Images are labelled with their capture coordinates and panorama IDs. Scripts for extending the dataset (i.e. fetching additional images)… See the full description on the dataset page: https://huggingface.co/datasets/everettshen/StreetView360AtoZ.imagetext-to-image10K<n<100K6 likes1.8k downloads2y agoHugging Face04FaroukMoc2 /jev-stage2-image-beans-pilot Beans: one natural question per image Open the corrected preview. natural_v4 is the recommended and default preview: 100 original images, 100 rows, one three-way condition-class Choice question per image. All targets come directly from the source labels column (34 angular leaf spot, 33 bean rust, 33 healthy). Original image bytes and source annotations are unchanged. Example question: “Which source-defined condition class describes the bean leaf?” Options: angular_leaf_spot… See the full description on the dataset page: https://huggingface.co/datasets/FaroukMoc2/jev-stage2-image-beans-pilot.imageimage-classificationn<1K0 likes802 downloads5d agoHugging Face05josefbednar /world-streetview-500k 🌍 World StreetView 500k World StreetView 500k is a large-scale computer vision dataset for visual geolocation estimation, spatial representation learning, and geographic scene understanding. It pairs ~2 million street-level images from ~500,000 unique locations worldwide with geographic coordinates, country labels, capture dates, and elevation data. Each training location is captured from 4 compass headings (0°, 90°, 180°, 270°) — ideal for training GeoGuessr-style geolocation… See the full description on the dataset page: https://huggingface.co/datasets/josefbednar/world-streetview-500k.imageimage-classification1M<n<10M1 likes791 downloads2mo agoHugging Face06Rapidata /human-style-preferences-images Rapidata Image Generation Preference Dataset This dataset was collected in ~4 Days using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation. Explore our latest model rankings on our website. If you get value from this dataset and would like to see more in the future, please consider liking it. Overview One of the largest human preference datasets for text-to-image models, this release contains over 1,200,000 human preference… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/human-style-preferences-images.imagetext-to-image10K<n<100K29 likes707 downloads2y agoHugging Face07FelicieGS /STHELAR_40x STHELAR dataset (40x) STHELAR (Spatial Transcriptomics and H&E histology for Large-scale Annotation Resource) is a multi-tissue dataset designed for developing models capable of predicting cell types directly from histological Hematoxylin & Eosin (H&E) whole slide images. It integrates high-resolution spatial transcriptomics data with histology, to provide detailed segmentation masks and cell-type annotations. Available dataset versions STHELAR_40x — 587,555 image… See the full description on the dataset page: https://huggingface.co/datasets/FelicieGS/STHELAR_40x.imageimage-segmentation100K<n<1M4 likes648 downloads6mo agoHugging Face08imageomics /STRI-Samples Dataset Card for Smithsonian Tropical Research Institute (STRI) Samples Dataset Summary Dorsal images of butterfly wings collected by Owen McMillan and members of his lab at the Smithsonian Tropical Research Institute. Full dataset will be 24,119 RGB images: Dorsal and Ventral images of separated wings. This sample contains 207 dorsal butterfly images used as part of the training data for Imageomics/butterfly_detection_yolo. Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/STRI-Samples.imageimage-classificationn<1K1 likes550 downloads9mo agoHugging Face09FelicieGS /STHELAR_20x STHELAR dataset (20x) STHELAR (Spatial Transcriptomics and H&E histology for Large-scale Annotation Resource) is a multi-tissue dataset designed for developing models capable of predicting cell types directly from histological Hematoxylin & Eosin (H&E) whole slide images. It integrates high-resolution spatial transcriptomics data with histology, to provide detailed segmentation masks and cell-type annotations. Available dataset versions STHELAR_40x — 587,555 image… See the full description on the dataset page: https://huggingface.co/datasets/FelicieGS/STHELAR_20x.imageimage-segmentation100K<n<1M4 likes424 downloads6mo agoHugging Face10stochastic /random_streetview_images_pano_v0.0.2 Dataset Card for panoramic street view images (v.0.0.2) Dataset Summary The random streetview images dataset are labeled, panoramic images scraped from randomstreetview.com. Each image shows a location accessible by Google Streetview that has been roughly combined to provide ~360 degree view of a single location. The dataset was designed with the intent to geolocate an image purely based on its visual content. Supported Tasks and Leaderboards None as of now!… See the full description on the dataset page: https://huggingface.co/datasets/stochastic/random_streetview_images_pano_v0.0.2.imageimage-classification10K<n<100K29 likes373 downloads4y agoHugging Face11paralaif /prl_crimson_maiden_style prl_crimson_maiden_style Dataset This dataset is used to train my lora Crimson Maiden from Civitai. imageimage-classificationn<1K1 likes320 downloads2y agoHugging Face12keremberke /painting-style-classification Dataset Labels ['Realism', 'Art_Nouveau_Modern', 'Analytical_Cubism', 'Cubism', 'Expressionism', 'Action_painting', 'Synthetic_Cubism', 'Symbolism', 'Ukiyo_e', 'Naive_Art_Primitivism', 'Post_Impressionism', 'Impressionism', 'Fauvism', 'Rococo', 'Minimalism', 'Mannerism_Late_Renaissance', 'Color_Field_Painting', 'High_Renaissance', 'Romanticism', 'Pop_Art', 'Contemporary_Realism', 'Baroque', 'New_Realism', 'Pointillism', 'Northern_Renaissance', 'Early_Renaissance'… See the full description on the dataset page: https://huggingface.co/datasets/keremberke/painting-style-classification.imageimage-classification1K<n<10K19 likes286 downloads4y agoHugging Face13TimSchneider42 /tactile-mnist-touch-starstruck-syn-single-t32-320x240Documentation is available at https://github.com/TimSchneider42/tactile-mnist/blob/main/doc/datasets.md#touch-datasets. imageimage-classification10K<n<100K0 likes281 downloads1y agoHugging Face14SABR22 /Canadian-streetview-cities Canadian Street View Cities Dataset Overview A street-view image dataset created to train and evaluate models for city-level image classification across major Canadian cities. Each entry includes an image and its corresponding city label. Purpose The dataset is intended for building models that recognize the Canadian city in which a street-view scene was captured. Data Source All images were collected from Mapillary, using geographic bounding boxes… See the full description on the dataset page: https://huggingface.co/datasets/SABR22/Canadian-streetview-cities.imageimage-classification100K<n<1M1 likes279 downloads10mo agoHugging Face15aeyxen /stem-diagrams STEM Diagrams 30,325 technical diagrams (block diagrams, schematics, flowcharts, architectures) extracted from arXiv papers across six engineering fields, each with a source attribution and a quality score. Built by an LLM-curated pipeline and used to show that a small frozen-feature classifier can replace the paid LLM labeling gate. Paper: Distilling an LLM Diagram-Curation Pipeline into Local Classifiers (Adnan Abbasi, Thothica, 2026) Code:… See the full description on the dataset page: https://huggingface.co/datasets/aeyxen/stem-diagrams.imageimage-classification10K<n<100K0 likes231 downloads2mo agoHugging Face16star092304 /typhoon-intensity-classification Typhoon - Image Classification Dataset This dataset comes from PTIT AI Challenge and is organized for a multi-class image classification task focusing on tropical cyclone (typhoon) intensity estimation. Dataset Structure The directory structure is organized as follows: train/ ├── images/ │ ├── image1.jpg │ └── ... └── annotations.csv (only present in the train folder) The public_test and private_test sets are used to evaluate and score the… See the full description on the dataset page: https://huggingface.co/datasets/star092304/typhoon-intensity-classification.imageimage-classification1K<n<10K1 likes185 downloads4mo agoHugging Face17Qin2006 /Strawberry-MM-Straw5 Dataset Card for Strawberry Disease Multimodal Dataset Dataset Description This is a multimodal dataset for strawberry disease detection, which contains strawberry image data, corresponding environmental parameters (air temperature, air humidity, soil moisture) and strawberry variety information. It can be used to study the correlation between environmental factors and strawberry disease occurrence, as well as multimodal fusion disease detection algorithms.… See the full description on the dataset page: https://huggingface.co/datasets/Qin2006/Strawberry-MM-Straw5.imageobject-detection1K<n<10K3 likes173 downloads5mo agoHugging Face18strangerguardhf /NSFW-MultiDomain-Classificationgated NSFW_MultiDomain The NSFW_MultiDomain dataset is a curated image classification dataset focused on multi-domain adult content recognition. It consists of 5 distinct categories aimed at facilitating the development of robust NSFW (Not Safe For Work) image classification models. This dataset enables training and benchmarking of models that can distinguish between subtle variations in explicit and non-explicit content across artistic, animated, and real-world imagery.… See the full description on the dataset page: https://huggingface.co/datasets/strangerguardhf/NSFW-MultiDomain-Classification.imageimage-classification10K<n<100K25 likes167 downloads1y agoHugging Face19klotz11 /plism-dataset-tiles-st Stain-Transferred PLISM Dataset Dataset Overview The Stain-Transferred PLISM dataset is a synthetic variant of the PLISM dataset tiles provided by Filiot et al. (2025), which is based on the original PLISM-wsi dataset by Ochi et al. (2024). This dataset isolates global color-level variations (staining profiles) from localized morphological and scanner-specific hardware artifacts. It is specifically designed to evaluate and robustify computational pathology… See the full description on the dataset page: https://huggingface.co/datasets/klotz11/plism-dataset-tiles-st.imageimage-feature-extraction1M<n<10M0 likes159 downloads2mo agoHugging Face20stevelohwc /pokemon_card_image_for_authenticity_classification Pokemon Card Image for Authenticity Classification This dataset contains front/back images of Pokemon cards for authenticity experiments. Dataset structure Images/: all image files (.jpeg) Images/metadata.jsonl: metadata used by Hugging Face imagefolder labels.csv: flat label file with the same rows as metadata Columns image: image object loaded from file id: image filename (unique id) side: card side (0 = front, 1 = back) labels: authenticity label (1 =… See the full description on the dataset page: https://huggingface.co/datasets/stevelohwc/pokemon_card_image_for_authenticity_classification.imageimage-classificationn<1K0 likes109 downloads7mo agoHugging Face21Project-AgML /strawberry_growth_detection Strawberry Growth Detection A dataset for detection of strawberry growth stages. The dataset contains 1,477 images with 3,997 bounding box annotations across 7 categories. The dataset also contains ground truth data related to the size of the strawberries from tagged leaves, as well as a decimal based growth stage. This dataset is indexed on https://project-agml.github.io/ as part of the AgML python library. Citation @article{yang2024predicting… See the full description on the dataset page: https://huggingface.co/datasets/Project-AgML/strawberry_growth_detection.imageobject-detection1K<n<10K0 likes100 downloads1mo agoHugging Face22Sadhana-24 /StreetView-Image-Dataset-10K Urban Streetscape Dataset for Vision Language Models A curated subset of 10,000 street view images with 25 essential features optimized for training vision language models on urban environment analysis tasks. Dataset Description This dataset contains street view imagery paired with comprehensive annotations covering infrastructure characteristics, visual perception metrics, environmental context, and semantic segmentation data. This comprehensive dataset represents a… See the full description on the dataset page: https://huggingface.co/datasets/Sadhana-24/StreetView-Image-Dataset-10K.imageimage-classification1K<n<10K8 likes95 downloads1y agoHugging Face23StephanAkkerman /stock-charts Stock Charts This dataset is a collection of a sample of images from tweets that I scraped using my Discord bot that keeps track of financial influencers on Twitter. The data consists of images that were part of tweets that mentioned a stock. This dataset can be used for a wide variety of tasks, such as image classification or feature extraction. FinTwit Charts Collection This dataset is part of a larger collection of datasets, scraped from Twitter and labeled by a… See the full description on the dataset page: https://huggingface.co/datasets/StephanAkkerman/stock-charts.imageimage-classification1K<n<10K11 likes72 downloads2y agoHugging Face24Reubencf /streetview-global StreetView Global A globally-sampled street-view image dataset with rich scene annotations and visual question-answer pairs. All photographs are sourced from Mapillary, the open street-level imagery platform, via its public image API. Each example pairs a street-level photograph with geographic metadata (latitude, longitude, compass, capture time, region), a free-form scene description, structured scene classification (setting, weather, time of day, road type, infrastructure), and… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/streetview-global.imageimage-to-text10K<n<100K5 likes68 downloads6mo agoHugging Face25ArinRoths /StairvsNonStair_Dataset 24-679 (Fall 2026): Stairs and Non-Stair Images ArinRoths/StairvsNonStair_Dataset Photos of stairs and non-stairs scenes, prepared as square RGB images with multiple separately generated training variants. The goal of this dataset is to classify whether or not stairs are present in an image. Source and task The original dataset contains 32 images that I collected and organized into stairs and non_stairs folders. There are 16 original stairs images and 16 original… See the full description on the dataset page: https://huggingface.co/datasets/ArinRoths/StairvsNonStair_Dataset.imageimage-classificationn<1K0 likes67 downloads7d agoHugging Face26HeyDonto /sd2-staged-foreign-objects SD2 staged-laboratory foreign-object frames (heydonto) 97 frames · 163 frame-level annotation rows · 8 staged laboratory takes · CC BY 4.0 This dataset discloses and carries the SD2 own-footage portion of the training data of the ORena SAVE FOCUS challenge entry's FRAME-track component: 163 of that component's 58,086 pooled training rows. The other sources of that corpus are not part of this dataset. What is in it frames/ — 97 JPEG frames (filename = first 16 hex… See the full description on the dataset page: https://huggingface.co/datasets/HeyDonto/sd2-staged-foreign-objects.imagevisual-question-answeringn<1K0 likes66 downloads21d agoHugging Face27vinod18vin /StreetView360AtoZStreetView 360X is a dataset containing 6342 360 degree equirectangular street view images randomly sampled and downloaded from Google Street View. It is published as part of the paper "StreetView360X: A Location-Conditioned Latent Diffusion Model for Generating Equirectangular 360 Degree Street Views" (Princeton COS Senior Independent Work by Everett Shen). Images are labelled with their capture coordinates and panorama IDs. Scripts for extending the dataset (i.e. fetching additional images)… See the full description on the dataset page: https://huggingface.co/datasets/vinod18vin/StreetView360AtoZ.imagetext-to-image1K<n<10K0 likes65 downloads4mo agoHugging Face28treborDev /hmdb51-pick-run-stand HMDB51 — Pick / Run / Stand (frames procesados) Subconjunto procesado del dataset HMDB51 para un caso de uso de clasificación de productividad de empleados en almacén mediante visión por computadora: distinguir entre trabajador activo (recogiendo / corriendo) e inactivo/pausado (parado). Clases Clase HMDB51 Etiqueta de negocio pick Activo (Pick Up) run Activo (Running) stand Inactivo/Pausado (Standing) Estadísticas 492 videos… See the full description on the dataset page: https://huggingface.co/datasets/treborDev/hmdb51-pick-run-stand.imageimage-classification1K<n<10K0 likes60 downloads3mo agoHugging Face29msolarz /StatMapCorpus StatMapCorpus v1 23,549 English-language statistical maps identified in MapPool, annotated for cartographic method by vision-language models. This repository is a mirror. The citable source of record is the deposit in Dane Badawcze UW (University of Warsaw, ICM): https://doi.org/10.58132/FSJGSP, version 1.0. The nine files in deposit/ are byte-identical to that release; deposit/schema.json carries their sha256 sums so you can verify this yourself. Cite the DOI, not this URL.… See the full description on the dataset page: https://huggingface.co/datasets/msolarz/StatMapCorpus.tabularimage-classification10K<n<100K0 likes58 downloads1mo agoHugging Face30StephanAkkerman /crypto-charts Crypto Charts This dataset is a collection of a sample of images from tweets that I scraped using my Discord bot that keeps track of financial influencers on Twitter. The data consists mainly of images that are cryptocurrency charts. This dataset can be used for a wide variety of tasks, such as image classification or feature extraction. FinTwit Charts Collection This dataset is part of a larger collection of datasets, scraped from Twitter and labeled by a human (me).… See the full description on the dataset page: https://huggingface.co/datasets/StephanAkkerman/crypto-charts.imageimage-classification1K<n<10K3 likes57 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.