datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
road-images-and-embeddings
Norwegian Road Images with Embeddings (Trondheim Area)
A dataset of 34,908 road images from the Trondheim region of Norway (~40km radius), captured by Statens vegvesen (Norwegian Public Roads Administration) in 2025. Each image is paired with rich geospatial metadata, nearest address information, and a 3072-dimensional image embedding from Google's gemini-embedding-2-preview model.
Dataset Structure
Each example contains:
Field
Type
Description
image
Image… See the full description on the dataset page: https://huggingface.co/datasets/thomasht86/road-images-and-embeddings.RoadBench
RoadBench
RoadBench is a benchmark for evaluating the fine-grained spatial understanding and reasoning
capabilities of multimodal large language models (MLLMs) in urban scenarios. It comprises eight
tasks spanning bird's-eye-view (BEV, satellite) and first-person-view (FPV, in-vehicle camera)
imagery, including lane counting, lane designation recognition, road network correction, road
type classification, and two cross-view tasks.
⚠️ Important: how BEV satellite… See the full description on the dataset page: https://huggingface.co/datasets/tsinghua-fib-lab/RoadBench.road-issues-detection-dataset
Road Issues Detection Dataset
Dataset Summary
This comprehensive dataset contains 9,660 high-resolution RGB images categorized for road infrastructure issues detection. The dataset focuses on identifying critical urban infrastructure problems including potholes, damaged roads, broken road signs, illegal parking violations, and environmental cleanliness issues. It has been specifically organized and curated for computer vision and machine learning applications in smart… See the full description on the dataset page: https://huggingface.co/datasets/Programmer-RD-AI/road-issues-detection-dataset.Road_Line_Marking_Dataset
Road Line and Marking Segmentation Dataset (RLMD)
This repository contains dataset and additional information for paper RLMD: A Dataset for Road Marking Segmentation.
Introduction
RLMD is a road line and marking semantic segmentation dataset containing 2137 driving scene images and annotations. The annotations is manually annotated with 25 categories and saved in polygon mask format. Information about the categories is shown bellow, or you can download the [csv].… See the full description on the dataset page: https://huggingface.co/datasets/veetinator/Road_Line_Marking_Dataset.Brazilian_Road_Signs_Dataset
Brazilian Road Signs Dataset
This dataset contains high-quality images of Brazilian road and traffic signs collected from various urban and rural environments. It supports AI research in computer vision, object detection, and autonomous driving systems adapted to Brazil’s signage standards and language.
Contact
For queries or collaborations related to this dataset, contact:
anoushka@kgen.io
abhishek.vadapalli@kgen.io
Supported Tasks
Task Categories:… See the full description on the dataset page: https://huggingface.co/datasets/HumynLabs/Brazilian_Road_Signs_Dataset.road-safety-vulnerable-road-users
Road Safety & Vulnerable Road Users Visual Dataset
Rows: 484
Dataset Description
Road Safety & Vulnerable Road Users Visual Dataset is a Global street-level imagery and geospatial computer vision dataset for image classification, visual search, and mobility research. The labeled features in the dataset are Crosswalk, Cyclist, School Bus, Traffic Cone, and Wheelchair, and each row records a matched visual feature label derived from the Outerview internal visual… See the full description on the dataset page: https://huggingface.co/datasets/Outerview/road-safety-vulnerable-road-users.IMDB-Face-Recognition
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/silk-road/IMDB-Face-Recognition.wildlife-near-roads-cities
Wildlife Near Roads & Cities Visual Dataset
Rows: 30,743
Dataset Description
Wildlife Near Roads & Cities Visual Dataset is a Global wildlife image dataset built from street-level imagery and geospatial computer vision records of animals observed near roads and cities. The labels are produced from query-driven visual matching in the Outerview internal visual search pipeline, using feature targets for Deer, Fox, Raccoon, Bird, Coyote, and Turtle. Included files are… See the full description on the dataset page: https://huggingface.co/datasets/Outerview/wildlife-near-roads-cities.global-road-sign-index
Outerview Global Road Signs Index
A large-scale geospatial index of road signs and traffic signage with latitude and longitude.
This dataset is part of Outerview’s mission to organize the world’s physical infrastructure and make it searchable, measurable, and continuously updated.
🌍 Overview
Feature: Road signs and traffic signage
Scope: Global
Entries: 15,000
Total Index: Millions of locations (full system)
Formats: Parquet / CSV / GeoJSON
This index… See the full description on the dataset page: https://huggingface.co/datasets/Outerview/global-road-sign-index.classified_fr_road_signs
France road signs classification dataset
This dataset contains a total of 66000+ detected road signs from the Panoramax street level pictures.
The detection model used is available at https://huggingface.co/Panoramax/detect_face_plate_sign
250+ classes of road signs have been created, each matching a sign official type:
Axx signs = danger or warning
Bxx signs = restrictions / forbiden
Cxx signs = information
CExx signs = touristic information
etc.
Check… See the full description on the dataset page: https://huggingface.co/datasets/Panoramax/classified_fr_road_signs.classified_de_road_signsThis dataset has been created using Panoramax pictures from Germany on which the https://huggingface.co/Panoramax/detect_face_plate_sign model has been used to detect road road signs and crop them.
It contain 20000+ photos of DE road signs in 130+ classes.
Additional "bad" or "other" classes contain non road signs (detection false positives or not yet classed signs).
For some classes, additionnal road signs have been added coming from european countries using similar signs.
The file names are… See the full description on the dataset page: https://huggingface.co/datasets/Panoramax/classified_de_road_signs.classified_nl_road_signsThis dataset has been created using Panoramax pictures from the Netherlands on which the https://huggingface.co/Panoramax/detect_face_plate_sign model has been used to detect road road signs and crop them.
It contain 24000+ photos of NL road signs in 150+ classes.
Additional "bad" or "other" classes contain non road signs (detection false positives or not yet classed signs).
For some classes, additionnal road signs have been added coming from european countries using similar signs.
The file… See the full description on the dataset page: https://huggingface.co/datasets/Panoramax/classified_nl_road_signs.classified_be_road_signsThis dataset has been created using Panoramax pictures from Belgium on which the https://huggingface.co/Panoramax/detect_face_plate_sign model has been used to detect road road signs and crop them.
It contain 24000+ photos of BE road signs in 140+ classes.
Additional "bad" or "other" classes contain non road signs (detection false positives or not yet classed signs).
For some classes, additionnal road signs have been added coming from european countries using similar signs.
The file names are… See the full description on the dataset page: https://huggingface.co/datasets/Panoramax/classified_be_road_signs.
