ud-synthetic/dutch-passports
Disclaimer: All passport images and associated data in this dataset are synthetically generated and do not correspond to real individuals. Any names, numbers, or personal details are fictional and used solely for research and development purposes. Introduction - Netherlands The Synthetic Netherlands Passports Dataset gathers more than 1,000 AI-generated passport images crafted for training OCR and computer vision systems on identity documents. Each record is fully synthetic… See the full description on the dataset page: https://huggingface.co/datasets/ud-synthetic/dutch-passports.
Disclaimer: All passport images and associated data in this dataset are synthetically generated and do not correspond to real individuals. Any names, numbers, or personal details are fictional and used solely for research and development purposes.
Introduction - Netherlands
The Synthetic Netherlands Passports Dataset gathers more than 1,000 AI-generated passport images crafted for training OCR and computer vision systems on identity documents. Each record is fully synthetic, so the collection holds no real personal data or biometric details — a privacy-compliant base for building identity verification, KYC, and fraud-detection pipelines. - [Get the data](https://unidata.pro/datasets/synthetic-passports/?utm_source=huggingface-synthetic&utm_medium=referral&utm_campaign=synthetic-netherlands-passports)
The dataset is scalable on demand — further images and metadata can be produced to match your specifications, typically within one week.
Each image is captured either on a clean white background or against a mix of realistic surroundings — desks, walls, and other everyday surfaces — so models adapt more reliably to production-grade conditions.
Coverage spans 50+ countries (including Belgium, Germany, Luxembourg, Austria, Denmark, and more). Submit a request via [the website](https://unidata.pro/datasets/synthetic-passports/?utm_source=huggingface-synthetic&utm_medium=referral&utm_campaign=synthetic-netherlands-passports) to learn more.
Every image is paired with detailed structured metadata describing the document's personal fields — passport number, full name, signature, date of birth, sex, place of birth, issuing authority, nationality, document type, and the machine-readable zone (MRZ) — plus technical attributes such as resolution and category.
Dataset general info
The dataset comprises 1,000+ synthetic passport images, each tied to a complete identity-style record and structured annotations. Additional samples can be produced on request within one week, and the imagery covers both clean white scenes and varied background environments.
Metadata fields include:
Use cases - Netherlands
Training document verification systems
Banks, fintech operators, and border-control teams can use this dataset to train models that authenticate identity documents across a wide range of real-world conditions. Structured metadata supports accurate field extraction and validation, while the diversity of backgrounds strengthens robustness once models reach production.
Building fraud detection pipelines
Compliance and security teams can apply these synthetic samples to compile high-quality training corpora without using real travel documents. Complete passport field coverage and full MRZ strings allow both standard records and anomalous patterns to be modelled and tested with confidence.
FAQ
Is this real-world or synthetic data?
All images are AI-generated and contain no biometric data or personal information tied to real individuals.
Can I request a custom dataset size?
Yes — the dataset is scalable, and additional samples can be generated based on your requirements within one week.
Can I request country-specific data?
Yes — support for 50+ countries is available. Please submit a request to get detailed coverage and samples.
Can I request a sample before purchasing?
Yes — free samples are available so you can evaluate image quality, metadata structure, and variation coverage before committing.
How is the dataset delivered?
After purchase, the dataset is delivered via secure AWS cloud infrastructure compliant with ISO 27001 and ISO 27701.
