---
title: "Synthetic Printed Turkish Passports Dataset"
description: "It is a synthetic Turkish passports dataset containing 5,000 high-quality, AI-generated images. Labeled with detailed metadata - including passport ID, class, gender, and lighting -…"
url: "https://unidata.pro/datasets/synthetic-turkish-passports/"
date_modified: "2026-06-08T10:36:56+03:00"
language: "en-US"
---
It is a synthetic Turkish passports dataset containing 5,000 high-quality, AI-generated images. Labeled with detailed metadata - including passport ID, class, gender, and lighting - this dataset supports PII extraction, identity verification, and biometric recognition system training while maintaining strict data protection standards.

## Dataset Structure

### The Numbers Section

**Numbers list:** - **Number:** 5000 — **Text:** Images

### Tooltip Section

**Tooltip items:**

- **Name:** PII
- **Name:** Data generation
- **Name:** Security
- **Name:** Anti-spoofing
- **Name:** Computer Vision

### Dataset Info

**Table with data:**

|  |  |
| --- | --- |
| Description | Printed synthetic passport images for training ML models in PII extraction |
| Data types | Image |
| Tasks | OCR, Computer Vision |
| Total number of files | 5 000 |
| Number of files in a set | 96 (Angles - 3, Lighting - 4, Backgrounds - 4, Distances - 2) |
| Angles | 0°, 25°, 45° |
| Lighting | Natural-daylight, Office-LED, Warm-indoor, Dim-light |
| Backgrounds | Neutral wall, Textured desk, Outdoor pavement, Docs-on-docs |
| Distance | Close (80-90 % frame), Medium (50-60 %) |
| Labeling | Metadata (Passport ID, Sample ID, Class, Gender, Age Group, Angle, Distance, Category, Resolution, Camera, Light Condition, Background, Timestamp) |
| Gender | Male, Female |

**Media Slider:**

- **Image in the slider:** ![](https://unidata.pro/wp-content/uploads/2025/10/turkey-passport-example2.webp)
- **Image in the slider:** ![](https://unidata.pro/wp-content/uploads/2025/10/turkish-passport-scaled.webp)

**Link to the sample:** [Download sample](https://drive.google.com/drive/folders/1jwlQLHa4AEJv29U90hGRNhLvJ187XG3l?usp=sharing)

### Technical  characteristics

**Table with data:**

|  |  |
| --- | --- |
| Image Extensions | JPG |
| Data Type | generated |

**Source and collection methodology:** Data was AI-generated.

### Dataset Use Cases - слайдер

**Industry Cards:**

- **Industry:** Border Control and Security Systems — **Title:** Training Biometric Verification Models — **Text:** Synthetic Printed Turkish Passports Dataset supports border control technologies by providing synthetic Turkish passport images that help train biometric verification and identity recognition systems. The dataset enables accurate detection of personal data and document authenticity while adhering to security and privacy standards applied in international travel systems.
- **Industry:** Financial and Identity Verification Services — **Title:** Enhancing Document Authentication Algorithms — **Text:** Banks and digital service providers can use Turkish Passport Dataset to improve automated ID checks and prevent fraud. With diverse lighting, angle, and background variations, the dataset strengthens recognition algorithms that verify Turkish citizens’ documents during account creation and electronic transactions.
- **Industry:** AI Research and Model Benchmarking — **Title:** Developing OCR and Computer Vision Models — **Text:** Synthetic Printed Turkish Passports Dataset provides high-quality images for OCR and machine learning research. It helps scientists and engineers benchmark recognition systems that extract textual and biometric information from travel documents, improving data accuracy across multilingual and cross-border applications.
- **Industry:** Software Testing and Quality Assurance — **Title:** Simulating Real-World Scenarios in Document Processing — **Text:** Software teams use this dataset to test document recognition software under different visual conditions. Synthetic generation allows developers to safely train and evaluate models on realistic, privacy-compliant data that replicates various real-world passport layouts and security features.

### Fact

**FAQs Heading:** FAQs

**List of Questions:**

- **Question:** What should I consider before buying the dataset? — **Answer:** Before purchasing this Turkish passport dataset, review the technical details such as image resolution, lighting variations, and background diversity. Ensure it aligns with your OCR, computer vision, or document classification project requirements. Since it’s a synthetic dataset, it contains no personal information from real Turkish citizens.
- **Question:** What types of annotations are provided? — **Answer:** Each passport image comes with detailed metadata annotations, including passport ID, sample ID, angle, distance, background type, and light condition. These annotations support accurate document detection, image classification, and synthetic identity verification tasks.
- **Question:** Can I request a sample of the dataset before purchasing or downloading it? — **Answer:** Yes, you can request a free sample of the synthetic passport dataset to evaluate its quality and format. The sample includes a small number of Turkish passport images demonstrating different lighting and angle conditions, helping you determine if the dataset meets your training needs.
- **Question:** Is it possible to request a custom dataset? — **Answer:** Yes. Unidata provides custom dataset generation services, allowing you to define parameters such as lighting conditions, backgrounds, angles, or document types. This ensures the dataset perfectly fits your machine learning or OCR model training objectives.
- **Question:** What was the sources of data for this Turkish Passports Dataset? — **Answer:** This Unidata dataset was created through controlled collection or synthetic generation, not taken from open or government sources. In this case, all Turkish passport images were AI-generated to mimic real-world conditions while maintaining full data privacy and compliance.
- **Question:** How are Unidata datasets licensed? — **Answer:** Unidata datasets follow a dual-licensing model: free samples are available for testing, while full datasets are available only through purchase. This approach allows users to evaluate dataset quality before investing in complete training data packages.
- **Question:** How are Unidata datasets stored? — **Answer:** All datasets are securely stored on AWS cloud infrastructure, ensuring high availability, data integrity, and privacy compliance. The system aligns with ISO 27001 and ISO 27701 standards, offering a secure environment for managing and accessing biometric and synthetic image data.
- **Question:** How long does it take to receive the dataset? — **Answer:** Once you submit a dataset request, Unidata will contact you to confirm the details and complete the documentation. After signing and payment, the dataset will be delivered within 3–10 business days via secure cloud access.
- **Question:** How does synthetic data improve privacy and security in document AI projects? — **Answer:** Synthetic data removes the need to use real passports containing sensitive personal information during model development and testing. This enables safer AI experimentation, easier dataset sharing, and more secure development of OCR and identity verification technologies.

[Full list of this site's AI-readable pages](https://unidata.pro/llms.txt)
