datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Obstacle-Detection-Dataset-YOLO
ROD-Dataset: Real-Time Obstacle Detection for Smartphone-Based Assistive Vision
24,326-image, 25-class YOLO dataset for obstacle detection
This dataset is the data product of our Real-Time Obstacle Detection (ROD) project at Amirkabir University of Technology, Tehran. The project addresses two related public-safety problems on the city sidewalk: the limited situational awareness of people living with visual impairments, and the elevated collision and fall risk for pedestrians… See the full description on the dataset page: https://huggingface.co/datasets/Abtinzandi/Obstacle-Detection-Dataset-YOLO.jimei-fire-smoke-yolo-datasetmilitary-labeled-yolo
Military-Labeled YOLO Dataset (DVIDS sourced)
Multi-class military object detection in YOLO format. Source images pulled from
the Defense Visual Information Distribution Service (DVIDS)
public domain library; labeled via in-house Gemini-VLM-assisted pipeline with
human-in-the-loop correction.
Classes (12)
ID
Name
0
soldier
1
tank
2
apc
3
artillery
4
mlrs
5
military_truck
6
helicopter
7
aircraft
8
warship
9
missile_launcher
10
car… See the full description on the dataset page: https://huggingface.co/datasets/llama-farm/military-labeled-yolo.tft-set17-unit-detector-yolo
TFT Set17 Unit Detector YOLO
YOLO-format unit detector dataset for TFT Set 17 experiments.
This package combines:
synthetic board screenshots made from arena textures and modelviewer unit renders
clean multi-angle modelviewer unit reference images
The dataset is intended for training a single-class unit object detector.
Structure
images/train/*.jpg
images/val/*.jpg
labels/train/*.txt
labels/val/*.txt
data.yaml
classes.txt
manifest.json
Counts… See the full description on the dataset page: https://huggingface.co/datasets/Ashen0li/tft-set17-unit-detector-yolo.YOLOSoccer
YOLO Soccer
A repository of robot soccer data for YOLO training.
NUpbr data generated using https://github.com/NUbots/NUpbr. Noisy annotations. Includes HDRs from Nagoya, Montreal, Sydney, Bangkok, Bordeaux and the old NUbots lab.
Eindhoven and old lab data annotated using https://labelstud.io/.
TORSO-21 data is converted from https://github.com/bit-bots/TORSO_21_dataset.
Garbage_Classification_YOLONotice: train set include 80% of original dataset, test and val sets have 10%.
dental-panoramic-xray-yolo
Dental Panoramic X-Ray Detection Dataset (YOLO Format)
Combined dataset for dental pathology detection on panoramic radiographs, in YOLO format. Built for training liodon-ai/dental-panoramic-detector.
Classes
ID
Name
Description
0
caries
Dental caries and deep caries
1
periapical_lesion
Periapical / apical periodontitis
2
impacted_tooth
Impacted and wisdom teeth
Dataset Sources
Source
Images
Boxes
License
DENTEX
724
3… See the full description on the dataset page: https://huggingface.co/datasets/liodon-ai/dental-panoramic-xray-yolo.digitize-pid-yolo
Digitize-PID (Symbols only), YOLO format
Note: I am not the author of this dataset
An annotated synthetic dataset, Dataset-P&ID, of 500 P&IDs with incorporates different
types of noise and complex symbols. This dataset contains only the symbols, i.e., under
the object detection task.
Paliwal, S., Jain, A., Sharma, M., & Vig, L. (2021). Digitize-PID: Automatic Digitization
of Piping and Instrumentation Diagrams. ArXiv, abs/2109.03794.
Original data repository:… See the full description on the dataset page: https://huggingface.co/datasets/hamzas/digitize-pid-yolo.oral-yolo-dataset
Oral YOLO Lesion Detector
This model is a YOLO-based object detection model trained to detect oral lesion regions in smartphone-captured oral cavity images.
The model is intended for research, prototyping, and assistive screening workflows. It is not a medical device and must not be used as the sole basis for diagnosis, treatment decisions, or clinical triage.
Model Details
Model type: YOLO object detector
Task: Oral lesion detection / object detection
Input: Oral cavity… See the full description on the dataset page: https://huggingface.co/datasets/sach3v/oral-yolo-dataset.asset-yolo-dataset
Asset YOLO Dataset
Auto-annotated. 84 classes.
cat-dog-yolo-dataset
Cat and Dog Detection Dataset (YOLO Format)
This dataset is designed for training and evaluating object detection models, specifically YOLOv8, YOLOv10, or YOLO11, to identify cats and dogs in images.
Dataset Structure
The dataset follows the standard YOLO object detection format:
images/: Contains the raw images (.jpg, .png).
labels/: Contains the corresponding bounding box annotations in .txt files.
Annotation Format
Each label file contains annotations in… See the full description on the dataset page: https://huggingface.co/datasets/sankalpa1998/cat-dog-yolo-dataset.Drone-Orthomosaic-Vehicles-Yolo-annotation
Dataset Tailings Mining Vehicles & Instruments (High-Res Drone Imagery)
Dataset Summary
This dataset contains high-resolution aerial imagery focused on vehicle detection and geotechnical monitoring instruments within active mining environments (tailings dams). The data was acquired using a DJI Zenmuse P1 sensor at 120m altitude.
Photogrammetric Context
The images originate from large-scale georeferenced orthomosaics generated from bi-daily… See the full description on the dataset page: https://huggingface.co/datasets/titoruizh/Drone-Orthomosaic-Vehicles-Yolo-annotation.spscd_plus_buoy_yolo
YOLO Object Detection Dataset
SPSCD + SMD buoy only
Dataset Information
Number of Classes: 13
Format: YOLO (ultralytics)
Splits: train, validation, test
Classes
Small Craft
Small Fishing Boat
Small Passenger Ship
Fishing Trawler
Large Passenger Ship
Sailing Boat
Speed Craft
Motorboat
Pleasure Yacht
Medium Ferry
Large Ferry
High Speed Craft
Buoy
Dataset Structure
dataset/
├── dataset.yaml # Dataset configuration
├── train/
│ ├── images/… See the full description on the dataset page: https://huggingface.co/datasets/ARG-NCTU/spscd_plus_buoy_yolo.varroa-yolo-under-2m-wiouPothole-detection-Yolov8MIMIC_YOLO_prediction_cxrTaiwan-black-bear-yolo
台灣黑熊偵測資料集 | Taiwan Black Bear Detection Dataset
🐻 用 AI 守護台灣國寶 🇹🇼
📖 資料集簡介
本資料集包含 18,112 張影像,專門用於訓練 台灣黑熊(學名:Ursus thibetanus formosanus)的物件偵測模型。所有標註均採用 YOLO 格式,可直接用於 YOLOv5、YOLOv8 等主流物件偵測框架。
🎯 核心特色
🌍 語言: 不適用(電腦視覺資料集)
📋 任務類型: 物件偵測(Object Detection)
🐾 應用領域: 野生動物保育、瀕危物種監測
💾 標註格式: YOLO v5/v8 相容格式
📊 類別數量: 1 類(台灣黑熊)
✨ 支援的應用場景
🔍 物件偵測: 在影像中偵測並定位台灣黑熊
📹 野外監測: 自動化野生動物追蹤與監控
🌲 保育研究: 支援瀕危物種保護工作
🚨 預警系統: 人熊衝突預防與警示
📊 資料集結構… See the full description on the dataset page: https://huggingface.co/datasets/alix2t7/Taiwan-black-bear-yolo.Compcap-with-yolo-v2cots_yolo_dataset
🪸 CSIRO Crown-of-Thorns Starfish (COTS) Detection Dataset — YOLO Format
This dataset is a modified version of the CSIRO COTS and COTS Scars Dataset, originally released under the Creative Commons Attribution 4.0 License (CC BY 4.0).
The original dataset contains images and annotations for Crown-of-Thorns Starfish (COTS) and COTS scars, collected to support coral reef monitoring and control efforts on the Great Barrier Reef (GBR).
These starfish are coral predators, and their… See the full description on the dataset page: https://huggingface.co/datasets/eloise54/cots_yolo_dataset.RTTS_YOLOmppe-dataset-yolo-v2
MPPE Dataset YOLO v2
Dataset orientado a la detección automática de elementos de protección personal médico mediante modelos de visión artificial.
Incluye imágenes anotadas con bounding boxes en formato YOLO para tareas de detección de objetos.
Información general
Nombre: MPPE Dataset YOLO v2
Tarea: Detección de objetos
Formato de anotaciones: YOLO
Formato de imágenes: JPG
Total de imágenes: 2931
División entrenamiento: 2344 imágenes
División validación: 587… See the full description on the dataset page: https://huggingface.co/datasets/stormbreaker20/mppe-dataset-yolo-v2.robust-nearfield-yoloWithUmbCUSTOM_YOLO_DATASETObstacle-Detection-Dataset-YOLO
ROD-Dataset: Real-Time Obstacle Detection for Smartphone-Based Assistive Vision
24,326-image, 25-class YOLO dataset for obstacle detection
This dataset is the data product of our Real-Time Obstacle Detection (ROD) project at Amirkabir University of Technology, Tehran. The project addresses two related public-safety problems on the city sidewalk: the limited situational awareness of people living with visual impairments, and the elevated collision and fall risk for pedestrians… See the full description on the dataset page: https://huggingface.co/datasets/ShafinSI/Obstacle-Detection-Dataset-YOLO.cubicasa5k-yolo
CubiCase5K in YOLO format (walls / doors / windows)
Converted from CubiCase5K (CC BY-NC 4.0):
SVG Wall/Railing, Door, Window polygons rasterized on F1_scaled.png
and traced back to YOLO polygon labels. Swing arcs are ignored.
Class 0 wall
Class 1 door (opening polygon only)
Class 2 window
Splits: 90/5/5 train/val/test (random, seed 0). Labels are polygon format,
usable for both detection and instance segmentation in ultralytics.
Obstacle-Detection-Dataset-YOLO
ROD-Dataset: Real-Time Obstacle Detection for Smartphone-Based Assistive Vision
24,326-image, 25-class YOLO dataset for obstacle detection
This dataset is the data product of our Real-Time Obstacle Detection (ROD) project at Amirkabir University of Technology, Tehran. The project addresses two related public-safety problems on the city sidewalk: the limited situational awareness of people living with visual impairments, and the elevated collision and fall risk for pedestrians… See the full description on the dataset page: https://huggingface.co/datasets/ty-li/Obstacle-Detection-Dataset-YOLO.sleap_mice_hc_yolo_pose
mice_hc — Home-cage mice pose (SLEAP) converted to YOLO pose format with Annolid (https://github.com/healthonrails/annolid)
Dataset Summary
mice_hc is a two-animal pose dataset consisting of pairs of male and female white Swiss Webster mice recorded from an overhead home-cage view with light bedding. The animals are low contrast relative to background, which makes it a useful benchmark for robust pose estimation in challenging conditions.
This Hugging Face dataset… See the full description on the dataset page: https://huggingface.co/datasets/healthonrails/sleap_mice_hc_yolo_pose.snakeaid-yolov12-5000-bbox
SnakeAid YOLOv12 5000 BBox
Dataset Summary
This repository contains a YOLO-format SnakeAid object-detection dataset for snake detection experiments. It is organized as image/label pairs across train, valid, test splits and is intended for training or evaluating YOLO-family detectors, including the related SnakeAid Detect YOLOv12 checkpoints linked below.
Safety note: snake detection can be safety-critical in real-world use. Treat model outputs trained on this data as… See the full description on the dataset page: https://huggingface.co/datasets/the-khiem7/snakeaid-yolov12-5000-bbox.car-parts-segmentation-yolo
AutoInspect - Car Parts Dataset (Ultralytics YOLO segmentation)
YOLO-версия датасета с сегментацией деталей авто Car Parts Dataset.
Часть проекта AutoInspect (pipeline: view classification → car parts segmentation → damage segmentation).
Основан на датасете от HITL. Для парных деталей были прставлены тэги side (left/right) при помощи Supervisely App. Список таких деталей:
Headlight
Tail-light
Mirror
Front-window
Back-window
Front-door
Back-door
Front-wheel
Back-wheel
Fender… See the full description on the dataset page: https://huggingface.co/datasets/mitbersh/car-parts-segmentation-yolo.Hazard-v10i-yolov8
