datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
typed_digital_signatures
Typed Digital Signatures Dataset
This comprehensive dataset contains synthetic digital signatures rendered across 30 different Google Fonts, specifically selected for their handwriting and signature-style characteristics. Each font contributes unique stylistic elements, making this dataset ideal for robust signature analysis and font recognition tasks.
Dataset Overview
Total Fonts: 30 different Google Fonts
Images per Font: 3,000 signatures
Total Dataset Size:… See the full description on the dataset page: https://huggingface.co/datasets/Benjy/typed_digital_signatures.American-Sign-Language-MNIST
Dataset Card for ASL-MNIST
This is a FiftyOne dataset with 34,627 samples of American Sign Language (ASL) alphabet images, converted from the original Kaggle Sign Language MNIST dataset into a format optimized for computer vision workflows.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/American-Sign-Language-MNIST.Signature-Verification-Dataset
Multilingual Signature Verification Dataset
Dataset Summary
The Multilingual Signature Verification Dataset is a curated collection of handwritten signatures designed for offline signature verification and related computer vision tasks.
The dataset contains more than 7,000 signature images spanning three major writing systems:
Hindi
Bengali
English
The English portion includes samples from the well-known CEDAR Signature Dataset, while additional Hindi and… See the full description on the dataset page: https://huggingface.co/datasets/rakshitdabral/Signature-Verification-Dataset.Marathi-Sign-Language
Marathi Sign Language Detection Dataset Card
Dataset Description
The Marathi Sign Language Dataset is a comprehensive collection of images designed to facilitate the development and training of machine learning models for recognizing Marathi sign language gestures. This dataset includes 43 distinct classes, each representing a unique character in the Marathi sign language alphabet. With approximately 1.2k images per class, the dataset totals over 51k images, all uniformly… See the full description on the dataset page: https://huggingface.co/datasets/VinayHajare/Marathi-Sign-Language.Synset-Signset-Germany-GTSRB-Subset
Synset Signset Germany - GTSRB Subset
The GTSRB subset of the Synset Signset Germany dataset addresses the task of traffic sign recognition in Germany. It contains the same 43 traffic
sign classes as the well-known GTSRB dataset and thus represents its “synthetic twin”. For each class, the dataset
contains 500 images, accumulating to 21,500 independent images in total.
Website: synset.de/datasets/synset-signset-ger/
Paper: Sielemann, A., Loercher, L., Schumacher, M. L., Wolf, S.… See the full description on the dataset page: https://huggingface.co/datasets/FraunhoferIOSB/Synset-Signset-Germany-GTSRB-Subset.digital_signatures
Digital Signatures Dataset
This dataset contains unique synthetic digital signatures rendered in different fonts:
4,000 synthetic signatures in Rage font
4,000 synthetic signatures in Mistral font
2,000 synthetic signatures in Arial Unicode font
Purpose
For the development of models that can detect digital signatures in documentation using the publicly available Docusign® font styles.
Structure
The dataset is organized into three folders:
rage/ - Contains… See the full description on the dataset page: https://huggingface.co/datasets/Benjy/digital_signatures.Synset-Signset-Germany
Synset Signset Germany
The Synset Signset Germany dataset addresses the task of traffic sign recognition in Germany. It contains a total of 105,500 images of 211 different
German traffic sign classes, including newly published (2020) and thus comparatively rare traffic signs. The subset of the first 43 classes in the dataset aims to represent
a “synthetic twin” of the well-known GTSRB dataset.
Website: synset.de/datasets/synset-signset-ger/
Paper: Sielemann, A., Loercher, L.… See the full description on the dataset page: https://huggingface.co/datasets/FraunhoferIOSB/Synset-Signset-Germany.handwriting-signatures-datasetA dataset of handwritten signatures with text prompts, designed for LoRA fine-tuning of diffusion models to generate realistic personal signatures.
Handwritten Signatures Dataset (Processed)
This dataset contains preprocessed handwritten signatures designed for signature verification and LoRA fine-tuning of diffusion models (e.g., Stable Diffusion) for text-to-image tasks.
📌 Processing steps
Labeling – assigned using computer vision and AI models to group… See the full description on the dataset page: https://huggingface.co/datasets/LuisitoLuisito/handwriting-signatures-dataset.American-Sign-Language-MNIST
Dataset Card for ASL-MNIST
This is a FiftyOne dataset with 34,627 samples of American Sign Language (ASL) alphabet images, converted from the original Kaggle Sign Language MNIST dataset into a format optimized for computer vision workflows.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available… See the full description on the dataset page: https://huggingface.co/datasets/saifmughal23/American-Sign-Language-MNIST.Signature-Verification-Dataset
Multilingual Signature Verification Dataset
Dataset Summary
The Multilingual Signature Verification Dataset is a curated collection of handwritten signatures designed for offline signature verification and related computer vision tasks.
The dataset contains more than 7,000 signature images spanning three major writing systems:
Hindi
Bengali
English
The English portion includes samples from the well-known CEDAR Signature Dataset, while additional Hindi and… See the full description on the dataset page: https://huggingface.co/datasets/Peachyy2208/Signature-Verification-Dataset.Traffic_Sign_Recogntion_DatabaseTSRD (Traffic Sign Recognition Database) 是一个中国交通标志数据集,包含多种交通标志类别。数据集分为训练集和测试集:
训练集:包含约4170张图像
测试集:包含约1994张图像
类别数:约58个不同的交通标志类别
数据集格式为:
图像文件名;宽;高;x1;y1;x2;y2;类别;
包含全种类数据集 / 4方向指示牌数据集
handwriting-signatures-datasetA dataset of handwritten signatures with text prompts, designed for LoRA fine-tuning of diffusion models to generate realistic personal signatures.
Handwritten Signatures Dataset (Processed)
This dataset contains preprocessed handwritten signatures designed for signature verification and LoRA fine-tuning of diffusion models (e.g., Stable Diffusion) for text-to-image tasks.
📌 Processing steps
Labeling – assigned using computer vision and AI models to group signatures… See the full description on the dataset page: https://huggingface.co/datasets/sl4shuur/handwriting-signatures-dataset.UK_Traffic_Sign_Inspection_Datasetasl_sign_languages_alphabets_v03Pakistani-Sign-Language
Pakistan Sign Language (PSL) Gesture Dataset
A landmark-based gesture recognition dataset for Pakistan Sign Language (PSL), built to support real-time sign language translation for deaf and hard-of-hearing communities in Pakistan. This dataset covers PSL alphabets, words, and sentences, making it one of the few structured and publicly available PSL resources in existence.
How the Data Was Collected
Each gesture was recorded via webcam and processed using MediaPipe in… See the full description on the dataset page: https://huggingface.co/datasets/Bakhtyar12/Pakistani-Sign-Language.Indian_Traffic_Sign_Image_Dataset
Indian Traffic Sign Image Dataset (Sample)
⚠️ This is a free sample subset for evaluation purposes only.The full dataset (2,000+ HD images) is available for commercial licensing.Contact: sales@datacluster.ai · datacluster.ai
Dataset Summary
This dataset is an extremely challenging collection of original Indian traffic sign images, crowdsourced from over 400 urban and rural areas. Every image is manually reviewed and verified by computer vision professionals at… See the full description on the dataset page: https://huggingface.co/datasets/Dataclusterlabspvtltd/Indian_Traffic_Sign_Image_Dataset.handwriting-signatures-datasetA dataset of handwritten signatures with text prompts, designed for LoRA fine-tuning of diffusion models to generate realistic personal signatures.
Handwritten Signatures Dataset (Processed)
This dataset contains preprocessed handwritten signatures designed for signature verification and LoRA fine-tuning of diffusion models (e.g., Stable Diffusion) for text-to-image tasks.
📌 Processing steps
Labeling – assigned using computer vision and AI models to group signatures… See the full description on the dataset page: https://huggingface.co/datasets/jade-kai/handwriting-signatures-dataset.malaysian-sign-language-dataset-v1
Malaysian Sign Language Dataset (V1)
Dataset Description
This dataset contains 170,000+ samples of Malaysian Sign Language (BIM).
It has been processed using MediaPipe to extract skeleton keypoints (1662 features per frame) and is organized by class.
Total Classes: [Insert Number, e.g., 80]
Format: .npy (NumPy arrays) stored inside .zip files (one zip per class).
Features: Pose, Left Hand, Right Hand landmarks.
Frame Length: Normalized to 30 frames per sequence.… See the full description on the dataset page: https://huggingface.co/datasets/PishangShedappp/malaysian-sign-language-dataset-v1.Brazilian_Road_Signs_Dataset
Brazilian Road Signs Dataset
This dataset contains high-quality images of Brazilian road and traffic signs collected from various urban and rural environments. It supports AI research in computer vision, object detection, and autonomous driving systems adapted to Brazil’s signage standards and language.
Contact
For queries or collaborations related to this dataset, contact:
anoushka@kgen.io
abhishek.vadapalli@kgen.io
Supported Tasks
Task Categories:… See the full description on the dataset page: https://huggingface.co/datasets/HumynLabs/Brazilian_Road_Signs_Dataset.handwriting-signatures-datasetA dataset of handwritten signatures with text prompts, designed for LoRA fine-tuning of diffusion models to generate realistic personal signatures.
Handwritten Signatures Dataset (Processed)
This dataset contains preprocessed handwritten signatures designed for signature verification and LoRA fine-tuning of diffusion models (e.g., Stable Diffusion) for text-to-image tasks.
📌 Processing steps
Labeling – assigned using computer vision and AI models to group signatures… See the full description on the dataset page: https://huggingface.co/datasets/ebixhaferaj/handwriting-signatures-dataset.signatures
Signatures
1176 cropped handwritten signature images (PNG, RGBA with transparent background), tightly cropped to the ink with a small uniform margin.
Files are named sig_0001.png ... sig_1176.png.
retail-streetscape-storefront-signals
Retail Streetscape & Storefront Signals Visual Dataset
Rows: 73,704
Dataset Description
Retail Streetscape & Storefront Signals Visual Dataset is a global wildlife image dataset-style street-level imagery collection focused on retail frontage, placemaking, and urban commerce signals in real-world scenes. The labeled target in this dataset is the feature field, which captures matched visual feature labels for Storefront, Restaurant, Coffee Shop, Pharmacy, Outdoor… See the full description on the dataset page: https://huggingface.co/datasets/Outerview/retail-streetscape-storefront-signals.classified_fr_road_signs
France road signs classification dataset
This dataset contains a total of 66000+ detected road signs from the Panoramax street level pictures.
The detection model used is available at https://huggingface.co/Panoramax/detect_face_plate_sign
250+ classes of road signs have been created, each matching a sign official type:
Axx signs = danger or warning
Bxx signs = restrictions / forbiden
Cxx signs = information
CExx signs = touristic information
etc.
Check… See the full description on the dataset page: https://huggingface.co/datasets/Panoramax/classified_fr_road_signs.asl_sign_languages_alphabets_v02classified_nl_road_signsThis dataset has been created using Panoramax pictures from the Netherlands on which the https://huggingface.co/Panoramax/detect_face_plate_sign model has been used to detect road road signs and crop them.
It contain 24000+ photos of NL road signs in 150+ classes.
Additional "bad" or "other" classes contain non road signs (detection false positives or not yet classed signs).
For some classes, additionnal road signs have been added coming from european countries using similar signs.
The file… See the full description on the dataset page: https://huggingface.co/datasets/Panoramax/classified_nl_road_signs.classified_de_road_signsThis dataset has been created using Panoramax pictures from Germany on which the https://huggingface.co/Panoramax/detect_face_plate_sign model has been used to detect road road signs and crop them.
It contain 20000+ photos of DE road signs in 130+ classes.
Additional "bad" or "other" classes contain non road signs (detection false positives or not yet classed signs).
For some classes, additionnal road signs have been added coming from european countries using similar signs.
The file names are… See the full description on the dataset page: https://huggingface.co/datasets/Panoramax/classified_de_road_signs.global-road-sign-index
Outerview Global Road Signs Index
A large-scale geospatial index of road signs and traffic signage with latitude and longitude.
This dataset is part of Outerview’s mission to organize the world’s physical infrastructure and make it searchable, measurable, and continuously updated.
🌍 Overview
Feature: Road signs and traffic signage
Scope: Global
Entries: 15,000
Total Index: Millions of locations (full system)
Formats: Parquet / CSV / GeoJSON
This index… See the full description on the dataset page: https://huggingface.co/datasets/Outerview/global-road-sign-index.classified_be_road_signsThis dataset has been created using Panoramax pictures from Belgium on which the https://huggingface.co/Panoramax/detect_face_plate_sign model has been used to detect road road signs and crop them.
It contain 24000+ photos of BE road signs in 140+ classes.
Additional "bad" or "other" classes contain non road signs (detection false positives or not yet classed signs).
For some classes, additionnal road signs have been added coming from european countries using similar signs.
The file names are… See the full description on the dataset page: https://huggingface.co/datasets/Panoramax/classified_be_road_signs.quebec-traffic-signs
Quebec Traffic Signs Dataset
Dataset Description
The Quebec Traffic Signs Dataset is a specialized image dataset designed for the recognition and interpretation of road and traffic signs specific to the province of Quebec, Canada. This dataset aims to provide a comprehensive collection of signs, including regulatory, warning, and informational signs, with a particular focus on the unique bilingual (French/English) and complex parking/construction signage prevalent in… See the full description on the dataset page: https://huggingface.co/datasets/RDLTechworks/quebec-traffic-signs.Harbor-Signs-Augmented
Harbor Signs Augmented
This release contains synthetic descriptions and label explanations generated from the Harbor Signs Source examples.
Transformation provenance
The transformation stage used Beacon-Vision-3B. Its model record states the full license label Apache License 2.0. No other model or collection was used during transformation.
The source examples remain linked through stable sample keys; this card records the transformation term separately for the… See the full description on the dataset page: https://huggingface.co/datasets/SOTAagi2030/Harbor-Signs-Augmented.
