Colombia Passport OCR with the StructOCR Python SDK
Install the official Python SDK, extract validated Colombia passport fields, and add the same workflow to a FastAPI service.

Why Colombian Passport OCR is Difficult in Python
Building a reliable passport OCR pipeline from scratch is non-trivial for Colombian documents. The Colombian passport features a bilingual layout, combining Spanish and English text on the same data page. Spanish uses accented characters such as á, é, í, ó, ú, ü, and the distinctive ñ, which frequently cause misinterpretation in standard Python OCR libraries like Tesseract, leading to character corruption and field misalignment. Furthermore, the documents incorporate intricate security backgrounds, guilloche patterns, and the national coat of arms that introduce optical noise. Developing custom RegEx patterns in Python to parse the Machine Readable Zone (MRZ) while correcting these bilingual and accented-character alignment issues is highly brittle, often resulting in high manual review rates.
Enterprise-Grade Extraction with the StructOCR SDK
StructOCR simplifies document processing in your Python ecosystem by replacing complex pipelines with a single async API call. Our service leverages pre-trained Deep Learning models optimized specifically for Latin American identity documents and the complexities of Spanish/English bilingual typography. Our passport mrz ocr api automatically handles perspective correction, denoising, and glare removal. Instead of returning raw, unstructured text strings, the StructOCR Python SDK provides standardized JSON output with validated fields. This capability is crucial for applications managing international borders, as it eliminates the need for manual parsing, delivering production-ready data directly to your FastAPI, Django, or Flask applications. For the fastest Python integration, install the official StructOCR SDK and see the Python SDK documentation.
Production Use Cases
- Digital Onboarding (e-KYC): Reduce drop-off rates by pre-filling user data from Colombian Passports into your fintech or digital services apps in under 2 seconds.
- Travel & Aviation Apps: Seamlessly integrate with Python backends for automated check-in systems and border management at hubs like El Dorado International Airport (BOG) in Bogotá.
- Financial Compliance: Ensure strict compliance with the Superintendencia Financiera de Colombia and regional Anti-Money Laundering (AML) regulations by automatically and accurately verifying identity documents.
Live Demo: Passport scanner
No registration required. Upload a file to test the extraction.
Drop files here or click to browse
JPG · PNG · WebP · up to 500 files · max 4.5 MB each
Colombia Passport OCR with the Python SDK
The official Python SDK handles file encoding, API communication, and structured passport results. Use the SDK directly or open the FastAPI tab for a server endpoint. Keep your API key in STRUCTOCR_API_KEY.
Prerequisite: `pip install structocr`; set the server-side `STRUCTOCR_API_KEY` environment variable.
Prefer another stack? Open the Node.js SDK + Express integration.
import os
from structocr import StructOCR
client = StructOCR(api_key=os.environ["STRUCTOCR_API_KEY"])
def scan_colombia_passport():
image_path = "colombia_passport_sample.jpg"
try:
result = client.scan_passport(image_path)
if result.get("success"):
data = result["data"]
print("Colombia passport extraction successful")
print(f"Passport #: {data.get('passport_number')}")
print(f"Name: {data.get('given_names')} {data.get('surname')}")
print(f"Nationality: {data.get('nationality')}")
print(f"DOB: {data.get('date_of_birth')}")
else:
print(f"Extraction failed: {result.get('error')}")
except Exception as error:
print(f"SDK error: {error}")
if __name__ == "__main__":
scan_colombia_passport()Technical Specs
- •Latency: < 4s (Average)
- •Uptime: 99.9% SLA
- •Security: AES-256 Encryption & SOC2 Compliant
- •Input: JPG, PNG, WebP, PDF (Max 4.5MB)
- •Output: JSON (Structured Data)
Key Features
- •Spanish Character Support: Accurately parses names and places containing accented characters such as á, é, í, ó, ú, ü, and ñ, without character corruption or alignment errors in your Python environment.
- •Visual Extraction (VIZ): Reliably extracts data directly from the visual inspection zone, bypassing intricate national security backgrounds and watermarks.
- •Date Normalization: Returns all dates (Birth, Issue, Expiry) in a standardized YYYY-MM-DD format, ready for Python date handling.
Sample JSON Output
The Python SDK returns a dictionary matching this normalized JSON structure.
{
"success": true,
"data": {
"type": "passport",
"country_code": "COL",
"nationality": "COL",
"passport_number": "AS123456",
"surname": "RODRÍGUEZ",
"given_names": "CARLOS ANDRÉS",
"sex": "M",
"date_of_birth": "1993-08-12",
"place_of_birth": "BOGOTÁ",
"date_of_issue": "2023-03-01",
"date_of_expiry": "2028-02-28",
"issuing_authority": "MINISTERIO DE RELACIONES EXTERIORES"
}
}Frequently Asked Questions
How does StructOCR compare to AWS Textract or Google Vision for Colombian documents?
Generic OCR services often struggle with the Spanish/English bilingual layout and special accented characters (á, é, í, ó, ú, ü, ñ) prevalent in Colombian passports, frequently misreading or omitting them. Furthermore, you remain responsible for writing Python parsing logic and validating MRZ checksums. StructOCR is a specialized API trained specifically on these Latin American documents, returning validated, labeled fields directly.
Do you store the uploaded images?
We do not store customer images. All data is processed in-memory (RAM) and is purged immediately after the API request is completed. We are a SOC2 compliant provider.
Where can I find the complete Python SDK documentation?
See the official Python SDK documentation for installation, authentication, supported methods, and FastAPI examples.
You May Also Like
Related tutorials, platform guides, and comparisons
Colombia Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Colombia passports.
Chile Passport OCR with Python SDK
Use the official StructOCR Python SDK and FastAPI to extract structured MRZ and VIZ data from Chile passports with a server-side integration.
Singapore Passport OCR with Python SDK
Use the official StructOCR Python SDK and FastAPI to extract structured MRZ and VIZ data from Singapore passports with a server-side integration.
Kazakhstan Passport OCR with Python SDK
Use the official StructOCR Python SDK and FastAPI to extract structured MRZ and VIZ data from Kazakhstan passports with a server-side integration.
Indonesia Passport OCR with Python SDK
Use the official StructOCR Python SDK and FastAPI to extract structured MRZ and VIZ data from Indonesia passports with a server-side integration.
South Korea Passport OCR with Python SDK
Use the official StructOCR Python SDK and FastAPI to extract structured MRZ and VIZ data from South Korea passports with a server-side integration.
Germany Passport OCR with Python SDK
Use the official StructOCR Python SDK and FastAPI to extract structured MRZ and VIZ data from Germany passports with a server-side integration.
France Passport OCR with Python SDK
Use the official StructOCR Python SDK and FastAPI to extract structured MRZ and VIZ data from France passports with a server-side integration.
Ghana Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Ghana passports.
Germany Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Germany passports.
Technical Comparisons & Integrations
Explore platform integrations and competitive analysis
Add Passport OCR to Your Bolt.new App in 5 Minutes
Step-by-step guide to integrating the StructOCR Passport scanner API into a Bolt.new app. Build travel, hospitality, and global KYC flows right in your browser.
Add Passport OCR using Cursor in 5 Minutes
Step-by-step guide to integrating the StructOCR Passport scanner API using the Cursor AI code editor. Build travel, hospitality, and global KYC flows rapidly.
Extracting ECR Status & File Numbers: Generic OCR vs. StructOCR Indian Passport API
Learn why generic text extractors fail to capture critical Indian KYC fields like ECR status and File Numbers, and how a specialized API structures this data automatically.
Heavy KYC SDKs vs. StructOCR Passport API for Mobile Apps
Discover why modern app developers are abandoning bulky on-device OCR SDKs in favor of lightweight, edge-deployed Passport APIs.
Add Passport OCR to Your Lovable.dev App in 5 Minutes
Step-by-step guide to integrating the StructOCR Passport scanner API into a Lovable.dev app through a secure server-side function for travel and KYC flows.
Add Passport OCR to Your Replit App in 5 Minutes
Step-by-step guide to integrating the StructOCR Passport scanner API into a Replit application. Build travel, hospitality, and global KYC flows using Replit Agent and Secrets.
From tutorial to production in 5 minutes.
You've seen the code. Now get your API key, grab your 200 free credits, and see it work with your own images. No credit card required.