datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AVM_Segmentation_train
Dataset Card for AVM (Around View Monitoring) Semantic Segmentation Dataset
This repository provides a FiftyOne-compatible version of the AVM semantic segmentation dataset for autonomous parking systems, with enhanced metadata and visualization capabilities.
This is a FiftyOne dataset with 6763 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/AVM_Segmentation_train.PInVerify
PInVerify Dataset
An offline embodied benchmark for Active Instance Verification (AIV).
Paper
arXiv:2605.30639
Code
github.com/Avalon-S/PInVerify
Project page
avalon-s.github.io/PInVerify
Venue
FMEA Workshop @ CVPR 2026 (Poster)
Overview
An agent that navigates to a target object does not always arrive at the right instance. Telling "white floral" from "white striped" takes a close look from more than one viewpoint, which is a separate… See the full description on the dataset page: https://huggingface.co/datasets/Avalon-S/PInVerify.food_not_food
Food vs Not Food Dataset (from Hugging Face ImageNet-1K)
This dataset is a binary classification subset derived from the Hugging Face imagenet-1k dataset. It is curated to support the task of distinguishing food images from non-food images.
📦 Dataset Overview
Source: imagenet-1k on Hugging Face Datasets
Classes:
food: 40 selected ImageNet classes representing food items (e.g., pizza, banana, hotdog)
not_food: 40 selected classes not related to food (e.g., car, clock… See the full description on the dataset page: https://huggingface.co/datasets/avnishs17/food_not_food.hass_avocadoThis dataset is a huggingface upload for https://data.mendeley.com/datasets/3xd9n945v8/1
Description from their website accessed on 2nd July 2025
This dataset consists of 14,710 labeled photographs of Hass avocados (Persea Americana Mill. cv Hass), resized to 800 x 800 pixels and saved in the .jpg format, designed to facilitate the development of deep learning models for predicting ripening stages and estimating shelf-life.
A total of 478 Hass avocados were acquired three days post-harvest and… See the full description on the dataset page: https://huggingface.co/datasets/c2p-cmd/hass_avocado.Nurisk-ICRA2026
Nurisk: VQA for Risk Assessment in Autonomous Driving
Nurisk is a visual question answering dataset focusing on risk assessment for autonomous driving. Each row contains:
image: a BEV image
question: a driving-related question
answer: the ground truth answer
Paper
NuRisk: A Visual Question Answering Dataset for Agent-Level Risk Assessment in Autonomous Driving — see the paper on arXiv:2509.25944 .
Framework
Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/TUM-AVS/Nurisk-ICRA2026.Nurisk
Nurisk: VQA for Risk Assessment in Autonomous Driving
Nurisk is a visual question answering dataset focusing on risk assessment for autonomous driving. Each row contains:
image: a BEV image
question: a driving-related question
answer: the ground truth answer
Paper
NuRisk: A Visual Question Answering Dataset for Agent-Level Risk Assessment in Autonomous Driving — see the paper on arXiv:2509.25944 .
Framework
Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/Yuan-avs/Nurisk.CaptchaOCR-500K
CaptchaOCR-500K
Dataset Summary
CaptchaOCR-500K is a large-scale CAPTCHA recognition dataset containing 500,000 CAPTCHA images with corresponding text labels.
The dataset is designed for training and evaluating Optical Character Recognition (OCR), CAPTCHA solving systems, image-to-text models, and computer vision models focused on text recognition.
Tasks
Optical Character Recognition (OCR)
CAPTCHA Recognition
Image-to-Text
Computer Vision
Text… See the full description on the dataset page: https://huggingface.co/datasets/AvinashRicky/CaptchaOCR-500K.Indian-plant-leaves-species
Indian Plant Leaves Species
Dataset Summary
This dataset consists of 592 high-resolution images of leaves from 12 different plant species. All images were captured using a mobile phone camera under natural lighting conditions. The dataset is intended for use in plant species classification, leaf recognition, and related computer vision tasks.
Supported Tasks and Leaderboards
Image Classification
Species Recognition
Transfer Learning for Plant Identification… See the full description on the dataset page: https://huggingface.co/datasets/avaishnav/Indian-plant-leaves-species.tripmatch-ai-dataset
TripMatch AI Dataset
A reproducible multimodal dataset for the TripMatch AI Final Project. It contains
10,000 synthetic text trip plans with a raw idea generated for every row by the
pretrained Hugging Face model google/flan-t5-small, plus 5,000 real street-view images
retained as extra multimodal work. The two configurations are separate so Dataset
Viewer can load each schema correctly.
Dataset statistics
Configuration
Rows
Main fields
Intended task… See the full description on the dataset page: https://huggingface.co/datasets/avihayamor/tripmatch-ai-dataset.civitai-top-nsfw-images-with-metadata
CivitAI Top NSFW Images Dataset
This dataset contains 6k+ top NSFW images from CivitAI filtered using top reactions. The dataset contains prompt & nsfw level metadata in prompts.json file. The nsfw levels are: Soft, Mature & X.
Original forum post:
https://diffused.to/Thread-CivitAI-Top-NSFW-Images-Dataset-6k-images
Dataset collection date
June 2025
Dataset structure:
├── 📂 images/
│ ├── 1.jpg
│ ├── 2.jpg
│ ├── 3.jpg
│ ├── ....
├──… See the full description on the dataset page: https://huggingface.co/datasets/Avanish11/civitai-top-nsfw-images-with-metadata.CW-AVIS
CW-AVIS
CW-AVIS is a CARLA-based benchmark dataset for cross-weather aerial visual
instance search. It pairs target descriptions and search start poses with
reference images captured under multiple weather and lighting conditions.
The dataset accompanies ReF-AVS.
Contents
2,470 PNG reference images
61 target definitions across six CARLA towns
Three search difficulties: easy, medium, and hard
Clear, cloudy, rainy, wet, night, sunset, and dust-storm conditions… See the full description on the dataset page: https://huggingface.co/datasets/Zoith/CW-AVIS.kawaii_chibi_avatar_dataset
Kawaii Chibi Avatar Dataset
This is the dataset used to train
Kawaii Chibi Avatar for Illustrious.
All images have a .txt file auto-tagged on Civitai.
All images were generated on SDXL using Kawaii Chibi Avatar for SDXL
License
License: CC BY 4.0
Attribution:
Kawaii Chibi Avatar Dataset © 2025 by Robb-0 is licensed under CC BY 4.0
avatar-the-last-airbender-tagged
Dataset Card for "avatar-the-last-airbender-tagged"
More Information needed
Medical_Prescription_Handwritten_Words
Medical Prescription Handwritten Words
This dataset contains images of individual handwritten medical words extracted from prescription notes. It is designed for training and evaluating handwriting recognition models in the healthcare domain.
Structure
images/: Contains 40+ handwritten word images (e.g., Amoxicillin.png, Cold.png, Tablet.png, 0.png, etc.)
data.csv: Maps each image file to its corresponding label (word)
Example Use Cases
OCR (Optical… See the full description on the dataset page: https://huggingface.co/datasets/avi-kai/Medical_Prescription_Handwritten_Words.AVAINT-IMGAVIBench
⚠️ Evaluation-only dataset. AVI-Bench is licensed under the
AVI-Bench Data Use Policy v1.0
(CC BY-ND 4.0 + Anti-Training Addendum).
Using this dataset, in whole or in part, to train, fine-tune,
distil, align, or otherwise update any machine-learning model is
expressly prohibited. Bulk redistribution and automated scraping
are also prohibited. Commercial evaluation and benchmarking are
permitted. See the full policy before downloading.
AVI-Bench
AVI-Bench: Toward… See the full description on the dataset page: https://huggingface.co/datasets/FudanCVL/AVIBench.AVA-Bench
AVA-Bench
Training dataset for the paper AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models (arXiv:2506.09082) accepted in CVPR 2026.
AVA-Bench is a diagnostic benchmark for evaluating Vision Foundation Models (VFMs) through Atomic Visual Abilities (AVAs): fundamental perceptual skills such as localization, counting, OCR, spatial understanding, depth estimation, color recognition, texture recognition, and fine-grained recognition.
AVA-Bench disentangls visual… See the full description on the dataset page: https://huggingface.co/datasets/act13/AVA-Bench.lookupjet-adsb-optical-tracking-jets-airplanes-aviation-samplesnull
Lookup-Jet: Multimodal Aviation Tracking Dataset
Overview
The Lookup-Jet Dataset is an industry-grade, sensor-fusion dataset designed for advanced computer vision and machine learning tasks. It combines high-resolution visual tracking (bounding boxes and polygon segmentations) of aircraft with synchronized ADS-B kinematics, environmental conditions, and astronomical data.
This dataset is pre-formatted for immediate deployment across standard ML frameworks… See the full description on the dataset page: https://huggingface.co/datasets/Lookupjet/lookupjet-adsb-optical-tracking-jets-airplanes-aviation-samples.
