---
title: "Synthetic Printed Brazilian Passports Dataset"
description: "Provides 5,000 high-resolution synthetic Brazilian passport images captured under diverse angles, lighting, and backgrounds. Designed as a synthetic ID dataset for OCR, computer vision, and…"
url: "https://unidata.pro/datasets/synthetic-printed-brazilian-passports/"
date_modified: "2026-02-16T22:39:07+03:00"
language: "en-US"
---
Provides 5,000 high-resolution synthetic Brazilian passport images captured under diverse angles, lighting, and backgrounds. Designed as a synthetic ID dataset for OCR, computer vision, and identity verification model training, it delivers realistic passport images without exposing real personal data or sensitive information.

## Dataset Structure

### The Numbers Section

**Numbered list:** - **Number:** 5 000 — **Text:** Images

### Tooltips Section

**Tooltip items:**

- **Name:** PII
- **Name:** Data generation
- **Name:** Security
- **Name:** Anti-spoofing
- **Name:** Computer Vision

### Dataset Information

**Table with data:**

| Characteristic | Data |
| --- | --- |
| Description | Printed synthetic passport images for training ML models in PII extraction |
| Data types | Image |
| Tasks | OCR, Computer Vision |
| Total number of files | 5 000 |
| Number of files in a set | 96 (Angles - 3, Lighting - 4, Backgrounds - 4, Distances - 2) |
| Angles | 0°, 25°, 45° |
| Lighting | Natural-daylight, Office-LED, Warm-indoor, Dim-light |
| Backgrounds | Neutral wall, Textured desk, Outdoor pavement, Docs-on-doc |
| Distance | Close (80-90 % frame), Medium (50-60 %) |
| Labeling | Metadata (Passport ID, Sample ID, Class, Country, Gender, Age Group, Angle, Distance, Category, Resolution, Camera, Light Condition, Background, Timestamp) |
| Gender | Male, Female |

**Media Slider:**

- **Image in the slider:** ![Synthetic Printed Brazilian Passports](https://unidata.pro/wp-content/uploads/2025/10/brazilian-passports-dataset-slider-1-scaled.webp)
- **Image in the slider:** ![Synthetic Printed Brazilian Passports](https://unidata.pro/wp-content/uploads/2025/10/brazilian-passports-dataset-slider--scaled.webp)

**Link to the sample:** [Download sample](https://drive.google.com/drive/folders/1Wln3wFy4OPhgQJ1wA2aGoduo_20cyu-x)

### Technical Specifications

**Table with data:**

| Characteristic | Data |
| --- | --- |
| Image Extensions | HEIC |
| Data Type | generated |

**Source and data collection methodology:** Source and collection methodology: Data was AI-generated.

### Dataset Use Cases - Slider

**Industry Cards:**

- **Industry:** Financial Services — **Title:** Fraud Detection and KYC Verification — **Text:** Banks and payment providers can use Brazilian passport datasets to strengthen KYC and AML processes. The synthetic passport images simulate real travel documents without exposing personal data, helping financial systems detect forged IDs, spot anomalies, and train secure verification models across different categories of identity documents.
- **Industry:** Government & Border Control — **Title:** Training Models for Travel Document Authentication — **Text:** This synthetic passport dataset supports government agencies and border security by providing access to diverse passport images for machine learning. The dataset includes variations in lighting, backgrounds, and angles, helping recognition systems identify fraudulent synthetic IDs and differentiate authentic Brazilian passports from manipulated documents or passports issued by other countries.
- **Industry:** Technology & AI Development — **Title:** OCR and Computer Vision Model Training — **Text:** AI developers rely on synthetic passport datasets for training OCR, document segmentation, and classification models. The dataset offers a broader range of annotated samples, including metadata on gender, age, and image conditions, enabling robust identity verification tools and improving accuracy in handling new passports or ID cards.
- **Industry:** Cybersecurity & Data Protection — **Title:** Testing Privacy-Safe Identity Solutions — **Text:** Since these passport images are fully synthetic, organizations can test biometric security, ID verification tools, and personal information extraction models without breaching privacy laws. This makes it a valuable alternative to real passport databases, offering high-quality training data for safeguarding online services, bank accounts, and sensitive personal information.

### Fact

**FAQs Heading:** FAQs

**List of Questions:**

- **Question:** What types of annotations are provided? — **Answer:** Annotations include structured metadata covering gender, age group, document category, resolution, angle, distance, lighting, and background type. This makes it suitable for training models on a broader range of scenarios.
- **Question:** What should I consider before buying this dataset? — **Answer:** Before purchasing the Brazilian passport dataset, consider whether you need synthetic ID images for OCR and computer vision tasks. This dataset is AI-generated, ensuring no personal data exposure, while still offering realistic training material for identity document processing.
- **Question:** Is it possible to request a custom dataset? — **Answer:** Yes. If your project requires synthetic IDs from other countries, additional categories, or specific annotation fields, Unidata can generate a tailored dataset.
- **Question:** Can I request a sample of the dataset before purchasing or downloading it? — **Answer:** Yes. Unidata provides free dataset samples for testing and evaluation, so you can validate image quality, metadata structure, and annotation consistency before committing to a full purchase.
- **Question:** Do Unidata datasets follow GDPR or other data privacy regulations? — **Answer:** Yes. All Unidata datasets are curated in compliance with GDPR and data protection regulations. Since this is a synthetic dataset, it avoids exposure of personal information while remaining lawful and ethical for use in machine learning and computer vision projects.
- **Question:** How are Unidata datasets stored? — **Answer:** Unidata securely stores all datasets on AWS cloud infrastructure, ensuring scalability and reliability. Storage complies with ISO standards, creating a secure environment for handling synthetic ID datasets.
- **Question:** How long does it take to receive the dataset? — **Answer:** After submitting a request, Unidata confirms details and prepares the documents. Following signing and payment, the synthetic Brazilian passport dataset is delivered within 3–10 business days.
- **Question:** Is this a real-world dataset or synthetic data? — **Answer:** This is a synthetic dataset generated using AI. While it replicates real Brazilian passports for research and OCR tasks, it does not contain data from passports issued to real people.
- **Question:** Why use a synthetic Brazilian passport dataset instead of real passport images? — **Answer:** A synthetic Brazilian passport dataset provides realistic document images without exposing real personal information, making it a privacy-friendly solution for AI development. It enables organizations to train OCR, document verification, and identity recognition models while reducing legal and compliance risks.

[Full list of this site's AI-readable pages](https://unidata.pro/llms.txt)
