datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
street2shopgoogle-streetview-images-by-country
Dataset Card for google streetview images by country
⚠️ There are still images that should be deleted, such as those with tags or those that didn't load correctly.
Dataset Structure
folder with the individual countries
images have the creation date and the map name in the file name.
Dataset Card Contact
use the community section
images per country
StreetViewHouseNumbers
Dataset Card for Street View House Numbers
The Street View House Numbers (SVHN) dataset is a large real-world image dataset used for developing machine learning and object recognition algorithms. It contains over 600,000 labeled images of house numbers taken from Google Street View.
The images are cropped to a fixed resolution of 32x32 pixels, centered around a single character but may contain some distractors at the sides.
SVHN is similar to the MNIST dataset but incorporates… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/StreetViewHouseNumbers.global-streetscapes
Global Streetscapes
Repository for the tabular portion of the Global Streetscapes dataset by the Urban Analytics Lab (UAL) at the National University of Singapore (NUS).
Content Breakdown
Global Streetscapes (74 GB)
├── data/ (49 GB)
│ ├── 21 CSV files with 346 unique features in total and 10M rows each (37 GB)
│ ├── parquet/ (12 GB) (New)
│ ├── 21 Parquet equivalents of the 21 CSV files (New)
│ ├── 1 combined Parquet file (New)
├── manual_labels/ (23… See the full description on the dataset page: https://huggingface.co/datasets/NUS-UAL/global-streetscapes.StreetView360AtoZStreetView 360X is a dataset containing 6342 360 degree equirectangular street view images randomly sampled and downloaded from Google Street View. It is published as part of the paper "StreetView360X: A Location-Conditioned Latent Diffusion Model for Generating Equirectangular 360 Degree Street Views" (Princeton COS Senior Independent Work by Everett Shen). Images are labelled with their capture coordinates and panorama IDs. Scripts for extending the dataset (i.e. fetching additional images)… See the full description on the dataset page: https://huggingface.co/datasets/everettshen/StreetView360AtoZ.delhi-street-conditionsworld-streetview-500k
🌍 World StreetView 500k
World StreetView 500k is a large-scale computer vision dataset for visual geolocation estimation, spatial representation learning, and geographic scene understanding.
It pairs ~2 million street-level images from ~500,000 unique locations worldwide with geographic coordinates, country labels, capture dates, and elevation data. Each training location is captured from 4 compass headings (0°, 90°, 180°, 270°) — ideal for training GeoGuessr-style geolocation… See the full description on the dataset page: https://huggingface.co/datasets/josefbednar/world-streetview-500k.5_synt_flux_street_selected_single_validated_1011streetview-graphsrandom_streetview_images_pano_v0.0.2
Dataset Card for panoramic street view images (v.0.0.2)
Dataset Summary
The random streetview images dataset are labeled, panoramic images scraped from randomstreetview.com. Each image shows a location
accessible by Google Streetview that has been roughly combined to provide ~360 degree view of a single location. The dataset was designed with the intent to geolocate an image purely based on its visual content.
Supported Tasks and Leaderboards
None as of now!… See the full description on the dataset page: https://huggingface.co/datasets/stochastic/random_streetview_images_pano_v0.0.2.global-streetscapes-parquet
Global Streetscapes
Repository for the tabular portion of the Global Streetscapes dataset by the Urban Analytics Lab (UAL) at the National University of Singapore (NUS).
Content Breakdown
Global Streetscapes (62+ GB)
├── data/ (37 GB)
│ ├── 21 CSV files with 346 unique features in total and 10M rows each
├── manual_labels/ (23 GB)
│ ├── train/
│ │ ├── 8 CSV files with manual labels for contextual attributes (training)
│ ├── test/
│ │ ├── 8 CSV files with… See the full description on the dataset page: https://huggingface.co/datasets/ClaireDons/global-streetscapes-parquet.michiyomi-tokyo-streetscape
michiyomi — Tokyo streetscape verbalization open data
English
Overview
michiyomi pairs coordinates with structured Japanese descriptions of physical streetscapes visible in public Mapillary imagery. A vision-language model (VLM) verbalized only what is visible in each image: no map, address, place name, facility name, statistics, or other external knowledge was injected. Release 2026-09-13-r1 contains 1,914,490 scenes covering all of Tokyo: the 23… See the full description on the dataset page: https://huggingface.co/datasets/finalvent/michiyomi-tokyo-streetscape.global-streetscapes
Global Streetscapes
Repository for the tabular portion of the Global Streetscapes dataset by the Urban Analytics Lab (UAL) at the National University of Singapore (NUS).
Content Breakdown
Global Streetscapes (74 GB)
├── data/ (49 GB)
│ ├── 21 CSV files with 346 unique features in total and 10M rows each (37 GB)
│ ├── parquet/ (12 GB) (New)
│ ├── 21 Parquet equivalents of the 21 CSV files (New)
│ ├── 1 combined Parquet file (New)
├── manual_labels/ (23… See the full description on the dataset page: https://huggingface.co/datasets/lune7723/global-streetscapes.streetview-acw-300kglobal-streetscapes
Global Streetscapes
Repository for the tabular portion of the Global Streetscapes dataset by the Urban Analytics Lab (UAL) at the National University of Singapore (NUS).
Content Breakdown
Global Streetscapes (74 GB)
├── data/ (49 GB)
│ ├── 21 CSV files with 346 unique features in total and 10M rows each (37 GB)
│ ├── parquet/ (12 GB) (New)
│ ├── 21 Parquet equivalents of the 21 CSV files (New)
│ ├── 1 combined Parquet file (New)
├── manual_labels/ (23… See the full description on the dataset page: https://huggingface.co/datasets/KevinAldrin/global-streetscapes.Canadian-streetview-cities
Canadian Street View Cities Dataset
Overview
A street-view image dataset created to train and evaluate models for city-level image classification across major Canadian cities. Each entry includes an image and its corresponding city label.
Purpose
The dataset is intended for building models that recognize the Canadian city in which a street-view scene was captured.
Data Source
All images were collected from Mapillary, using geographic bounding boxes… See the full description on the dataset page: https://huggingface.co/datasets/SABR22/Canadian-streetview-cities.RenderMatte-Human-Multi-Street-2K
RenderMatte Human Multi-Person Street 2K
A synthetic video-matting dataset for the multi-person case. Two or three rigged 3D human characters are
animated and rendered together in one Blender (Cycles) scene, then composited onto real street footage.
Every frame comes with its 16-bit alpha matte.
It is built on top of the single-person VideoMatting pipeline and its source-asset collection, and keeps the
same render and compositing settings wherever possible.
Research use only.… See the full description on the dataset page: https://huggingface.co/datasets/Deanshy1/RenderMatte-Human-Multi-Street-2K.prague-streetview-50kStreetviewLLM_Datastreetview-nyc-images5_synt_flux_street_selected_multi_v1_10225_synt_flux_street_selectedstreet-smart-road-signs
Street Smart: Road Sign Recognition
Dataset Summary
A public, viewer-ready educational challenge dataset. Host-only scoring data and hidden targets are excluded.
Splits
Split
Examples
Description
train
701
Labeled training data
test
176
Public inputs with withheld target labels or annotations
Data Fields
Field
Type
image
Image
image_id
string
width
int64
height
int64
objects.bbox… See the full description on the dataset page: https://huggingface.co/datasets/hoangbang/street-smart-road-signs.Street-garbage
Street-garbage
Описание
Street-garbage — публичный датасет для задачи детекции мусора на уличных изображениях городских сцен. Большинство изображений собрано в Москве и Санкт‑Петербурге; в набор включены дневные и ночные кадры, сцены с единичным и множественным мусором, а также кадры без мусора (чистые сцены) для снижения ложных срабатываний.
Размер и разбиение
train: 786 изображений
validation: 136 изображений
test: 144 изображения
всего: 1066 изображений… See the full description on the dataset page: https://huggingface.co/datasets/Shynuaa/Street-garbage.Street-videos
Dataset Card for Street Videos Dataset
This dataset contains a collection of street videos designed to support tasks such as video classification, object detection, and urban scene analysis. Each video captures various street environments, including urban settings, pedestrian areas, and traffic scenarios, to provide diverse training data for computer vision models.
Dataset Details
Dataset Description
This dataset contains videos of street scenes contributed by… See the full description on the dataset page: https://huggingface.co/datasets/HumynLabs/Street-videos.street-work
Street Work
This dataset is part of the Roboflow 100 benchmark, a diverse collection of 100 object detection datasets spanning 7 imagery domains.
Dataset Statistics
Split
Images
Train
611
Validation
175
Test
87
Total
873
Classes (11)
Cone
Coverall
Face_Shield
Gloves
Goggles
Head
Helmet
Mask
No glasses
No gloves
Person
Usage
With LibreYOLO
from libreyolo import LIBREYOLO
# Load a model
model =… See the full description on the dataset page: https://huggingface.co/datasets/LibreYOLO/street-work.Bangladeshi_Street_FoodStreetSignSet
Street Sign Set
High-Quality Traffic Sign Detection Dataset
📂 Dataset Overview
Street Sign Set is a comprehensive dataset designed for road sign detection in realistic contexts. It serves as the foundation for the StreetSignSense project, enabling robust detection in diverse environmental conditions.
The dataset is not perfectly balanced, reflecting the real-world frequency where some signs appear much more often than others.… See the full description on the dataset page: https://huggingface.co/datasets/AlessandroFerrante/StreetSignSet.StreetVision-10K
StreetVision-10K
Each sample contains:
A system prompt instructing the model to act as an OSINT/geospatial expert
A user message with a street-level photo and the instruction to determine coordinates
An assistant response with ground-truth coordinates in <direct_lon_lat_output>longitude,latitude</direct_lon_lat_output> format
Format
Each line is a JSON array of ChatML messages:
[
{"role": "system", "content": "..."},
{"role": "user", "content": [… See the full description on the dataset page: https://huggingface.co/datasets/mishl/StreetVision-10K.streetreview
Paper (HF): https://huggingface.co/papers/2508.11708
Repository: https://github.com/rsdmu/streetreview
StreetReview Dataset
Overview
StreetReview is a curated dataset designed to evaluate the inclusivity, accessibility, aesthetics, and practicality of urban streetscapes, particularly in a multicultural city context. Focused on Montréal, Canada, the dataset combines diverse demographic evaluations with rich metadata and street-view imagery. It aims to advance research… See the full description on the dataset page: https://huggingface.co/datasets/rsdmu/streetreview.
