datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Obstacle-Detection-Dataset-YOLO
ROD-Dataset: Real-Time Obstacle Detection for Smartphone-Based Assistive Vision
24,326-image, 25-class YOLO dataset for obstacle detection
This dataset is the data product of our Real-Time Obstacle Detection (ROD) project at Amirkabir University of Technology, Tehran. The project addresses two related public-safety problems on the city sidewalk: the limited situational awareness of people living with visual impairments, and the elevated collision and fall risk for pedestrians… See the full description on the dataset page: https://huggingface.co/datasets/Abtinzandi/Obstacle-Detection-Dataset-YOLO.jimei-fire-smoke-yolo-datasetkitchen_rack_combo_v2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 200,
"total_frames": 266498,
"total_tasks": 1,
"total_videos": 600,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:200"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/kitchen_rack_combo_v2.kitti-yolo11n-robustness-benchmark
KITTI YOLO11n Robustness & Adversarial Benchmark Suite
This dataset contains 649,425 benchmark samples evaluating the perception robustness of YOLO11n (Ultralytics YOLOv11 nano in original FP32 precision) on the official KITTI Object Detection train set (3,711 images) under 35 attack & corruption techniques across 5 severity levels.
?? Benchmark Leaderboard (mAP@0.5 Drop on YOLO11n)
Clean Baseline AP50: 0.3555
Evaluation Model: YOLO11n (Original weights:… See the full description on the dataset page: https://huggingface.co/datasets/VietPhong/kitti-yolo11n-robustness-benchmark.kitchen_rack_comboThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 196,
"total_frames": 258673,
"total_tasks": 1,
"total_videos": 588,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:196"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/kitchen_rack_combo.military-labeled-yolo
Military-Labeled YOLO Dataset (DVIDS sourced)
Multi-class military object detection in YOLO format. Source images pulled from
the Defense Visual Information Distribution Service (DVIDS)
public domain library; labeled via in-house Gemini-VLM-assisted pipeline with
human-in-the-loop correction.
Classes (12)
ID
Name
0
soldier
1
tank
2
apc
3
artillery
4
mlrs
5
military_truck
6
helicopter
7
aircraft
8
warship
9
missile_launcher
10
car… See the full description on the dataset page: https://huggingface.co/datasets/llama-farm/military-labeled-yolo.kitchen_rack_combo_v2_spoon_onlyThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 183,
"total_frames": 75265,
"total_tasks": 1,
"total_videos": 549,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:183"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/kitchen_rack_combo_v2_spoon_only.tft-set17-unit-detector-yolo
TFT Set17 Unit Detector YOLO
YOLO-format unit detector dataset for TFT Set 17 experiments.
This package combines:
synthetic board screenshots made from arena textures and modelviewer unit renders
clean multi-angle modelviewer unit reference images
The dataset is intended for training a single-class unit object detector.
Structure
images/train/*.jpg
images/val/*.jpg
labels/train/*.txt
labels/val/*.txt
data.yaml
classes.txt
manifest.json
Counts… See the full description on the dataset page: https://huggingface.co/datasets/Ashen0li/tft-set17-unit-detector-yolo.arabic_text_yolov8_V0.5YOLOSoccer
YOLO Soccer
A repository of robot soccer data for YOLO training.
NUpbr data generated using https://github.com/NUbots/NUpbr. Noisy annotations. Includes HDRs from Nagoya, Montreal, Sydney, Bangkok, Bordeaux and the old NUbots lab.
Eindhoven and old lab data annotated using https://labelstud.io/.
TORSO-21 data is converted from https://github.com/bit-bots/TORSO_21_dataset.
Garbage_Classification_YOLONotice: train set include 80% of original dataset, test and val sets have 10%.
dental-panoramic-xray-yolo
Dental Panoramic X-Ray Detection Dataset (YOLO Format)
Combined dataset for dental pathology detection on panoramic radiographs, in YOLO format. Built for training liodon-ai/dental-panoramic-detector.
Classes
ID
Name
Description
0
caries
Dental caries and deep caries
1
periapical_lesion
Periapical / apical periodontitis
2
impacted_tooth
Impacted and wisdom teeth
Dataset Sources
Source
Images
Boxes
License
DENTEX
724
3… See the full description on the dataset page: https://huggingface.co/datasets/liodon-ai/dental-panoramic-xray-yolo.vindr-png-yolo-rescaledigitize-pid-yolo
Digitize-PID (Symbols only), YOLO format
Note: I am not the author of this dataset
An annotated synthetic dataset, Dataset-P&ID, of 500 P&IDs with incorporates different
types of noise and complex symbols. This dataset contains only the symbols, i.e., under
the object detection task.
Paliwal, S., Jain, A., Sharma, M., & Vig, L. (2021). Digitize-PID: Automatic Digitization
of Piping and Instrumentation Diagrams. ArXiv, abs/2109.03794.
Original data repository:… See the full description on the dataset page: https://huggingface.co/datasets/hamzas/digitize-pid-yolo.Trucks-Detection-Yolov8
Trucks Detection - v1
This dataset was exported via roboflow.com on September 11, 2023 at 8:38 AM GMT
Roboflow is an end-to-end computer vision platform that helps you
collaborate with your team on computer vision projects
collect & organize images
understand and search unstructured image data
annotate, and create datasets
export, train, and deploy computer vision models
use active learning to improve your dataset over time
The dataset includes 746 images.
Trucks are annotated in… See the full description on the dataset page: https://huggingface.co/datasets/beethogedeon/Trucks-Detection-Yolov8.oral-yolo-dataset
Oral YOLO Lesion Detector
This model is a YOLO-based object detection model trained to detect oral lesion regions in smartphone-captured oral cavity images.
The model is intended for research, prototyping, and assistive screening workflows. It is not a medical device and must not be used as the sole basis for diagnosis, treatment decisions, or clinical triage.
Model Details
Model type: YOLO object detector
Task: Oral lesion detection / object detection
Input: Oral cavity… See the full description on the dataset page: https://huggingface.co/datasets/sach3v/oral-yolo-dataset.face-hands-YOLOv5yolo-baselines-no-mosaic-musgd-runstinyperson_mmdet_yolo_protocol_seed42_runsyolo-baselines-no-mosaic-runsolmo-3-preference-mix-deltas_reasoning-yolo_scottmix-DECON-multi-turnPothole-detection-Yolov8tinyperson_mmdet_yolo_protocol_seed43_runsDrone-Orthomosaic-Vehicles-Yolo-annotation
Dataset Tailings Mining Vehicles & Instruments (High-Res Drone Imagery)
Dataset Summary
This dataset contains high-resolution aerial imagery focused on vehicle detection and geotechnical monitoring instruments within active mining environments (tailings dams). The data was acquired using a DJI Zenmuse P1 sensor at 120m altitude.
Photogrammetric Context
The images originate from large-scale georeferenced orthomosaics generated from bi-daily… See the full description on the dataset page: https://huggingface.co/datasets/titoruizh/Drone-Orthomosaic-Vehicles-Yolo-annotation.asset-yolo-dataset
Asset YOLO Dataset
Auto-annotated. 84 classes.
cat-dog-yolo-dataset
Cat and Dog Detection Dataset (YOLO Format)
This dataset is designed for training and evaluating object detection models, specifically YOLOv8, YOLOv10, or YOLO11, to identify cats and dogs in images.
Dataset Structure
The dataset follows the standard YOLO object detection format:
images/: Contains the raw images (.jpg, .png).
labels/: Contains the corresponding bounding box annotations in .txt files.
Annotation Format
Each label file contains annotations in… See the full description on the dataset page: https://huggingface.co/datasets/sankalpa1998/cat-dog-yolo-dataset.varroa_mmdet_yolo_protocol_runsKu-Yolo-DataSet
Ku-YOLO Waste Classification Dataset
쓰레기 분류를 위한 YOLO 형식 데이터셋입니다. TACO 데이터셋을 기반으로 재활용품 객체 탐지를 위해 재구성되었습니다.
데이터셋 개요
총 이미지 수: 1,500개
총 어노테이션 수: 4,784개
이미지 형식: JPG
어노테이션 형식: YOLO TXT
클래스 정보
5개 대분류 (labels_5classes)
ID
클래스
영문
어노테이션 수
비율
0
플라스틱
Plastic
1,443
30.16%
1
비닐
Vinyl
845
17.66%
2
캔
Can
552
11.54%
3
유리
Glass
254
5.31%
4
종이
Paper
1,690
35.33%
60개 세부 분류 (labels)
플라스틱 병, 유리병, 음료수 캔, 종이컵, 비닐봉투 등 60개의 세부 카테고리로… See the full description on the dataset page: https://huggingface.co/datasets/hyeon2525/Ku-Yolo-DataSet.varroa-yolo-under-2m-wiouclean_up_table_v2.1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 194,
"total_frames": 148588,
"total_tasks": 5,
"total_videos": 582,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:194"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/clean_up_table_v2.1.
