Bulgaria Passport OCR with the StructOCR Node.js SDK
Install the official Node.js SDK, extract validated Bulgaria passport fields, and add the same workflow to a TypeScript Express service.

Why Bulgarian Passport OCR is Difficult in Node.js
Building a reliable passport OCR pipeline from scratch is non-trivial for Bulgarian documents. The Bulgarian passport features a bilingual layout, presenting personal data in Bulgarian Cyrillic alongside English. The coexistence of Cyrillic characters—including unique letters such as ъ, ь, ю, я, and the distinctive щ—with Latin script frequently confuses standard Node.js wrappers like Tesseract, causing character substitution errors (e.g., В/B, Р/P) and misaligned field extraction. Furthermore, the documents incorporate intricate security backgrounds, guilloche patterns, and the national coat of arms that introduce optical noise. Developing custom RegEx patterns in JavaScript to parse the Machine Readable Zone (MRZ) while correcting these script-mixing alignment issues is highly brittle, often resulting in high manual review rates.
Enterprise-Grade Extraction with the StructOCR SDK
StructOCR simplifies document processing in your JavaScript ecosystem by replacing complex pipelines with a single async API call. Our service leverages pre-trained Deep Learning models optimized specifically for Southeastern European identity documents and the complexities of Cyrillic/English bilingual typography. Our passport mrz ocr api automatically handles perspective correction, denoising, and glare removal. Instead of returning raw, unstructured text strings, the StructOCR Node.js SDK provides standardized JSON output with validated fields. This capability is crucial for applications managing international borders, as it eliminates the need for manual parsing, delivering production-ready data directly to your Express, NestJS, or Next.js applications. For the fastest Node.js integration, install the official StructOCR SDK and see the Node.js SDK documentation.
Production Use Cases
- Digital Onboarding (e-KYC): Reduce drop-off rates by pre-filling user data from Bulgarian Passports into your fintech or digital services apps in under 2 seconds.
- Travel & Aviation Apps: Seamlessly integrate with Node.js backends for automated check-in systems and border management at hubs like Sofia Airport (SOF).
- Financial Compliance: Ensure strict compliance with the Bulgarian National Bank (BNB) and European Anti-Money Laundering (AML) regulations by automatically and accurately verifying identity documents.
Live Demo: Passport scanner
No registration required. Upload a file to test the extraction.
Drop files here or click to browse
JPG · PNG · WebP · up to 500 files · max 4.5 MB each
Bulgaria Passport OCR with the Node.js SDK
The official Node.js SDK handles file buffers, API communication, and structured passport results. Use the SDK directly or open the TypeScript + Express tab for a server endpoint. Keep your API key in STRUCTOCR_API_KEY.
Prerequisite: `npm install structocr`; set the server-side `STRUCTOCR_API_KEY` environment variable.
Prefer another stack? Open the Python SDK + FastAPI integration.
const StructOCR = require('structocr');
// Initialize the client with your API key
const apiKey = process.env.STRUCTOCR_API_KEY;
if (!apiKey) throw new Error('STRUCTOCR_API_KEY is not configured');
const client = new StructOCR(apiKey);
async function scanBulgariaPassport() {
// Note: Supports JPG, PNG, WebP, PDF (Max 4.5MB)
const imagePath = './bulgaria_passport_sample.jpg';
try {
console.log(`Processing Bulgaria passport: ${imagePath}...`);
// The SDK handles file reading and the API call seamlessly
const result = await client.scanPassport(imagePath);
if (result.success && result.data) {
const data = result.data;
console.log('✅ Extraction Successful!');
// Basic Identity (MRZ + VIZ)
console.log(`Passport #: ${data.passport_number}`);
console.log(`Name: ${data.surname} ${data.given_names}`);
console.log(`Nation: ${data.nationality} (${data.country_code})`);
// Visual Zone Specifics
console.log(`Birth Place:${data.place_of_birth}`);
console.log(`Authority: ${data.issuing_authority}`);
// Dates
console.log(`DOB: ${data.date_of_birth} (${data.sex})`);
console.log(`Expiry: ${data.date_of_expiry}`);
} else {
console.error('❌ Extraction Failed:', result.error || 'Unknown Error');
}
} catch (error) {
console.error('An unexpected error occurred:', error.message);
}
}
scanBulgariaPassport();Technical Specs
- •Latency: < 4s (Average)
- •Uptime: 99.9% SLA
- •Security: AES-256 Encryption & SOC2 Compliant
- •Input: JPG, PNG, WebP, PDF (file path, Buffer, or Uint8Array; SDK converts to Base64 JSON)
- •Max File Size: 4.5MB
- •Output: JSON (Structured Data)
Key Features
- •Cyrillic-Latin Bilingual Support: Accurately parses names and descriptors across Bulgarian Cyrillic and English scripts, resolving confusable characters like В/B or Р/P and preserving unique letters like ъ, ь, ю, я, щ without alignment errors in your Node environment.
- •Visual Extraction (VIZ): Reliably extracts data directly from the visual inspection zone, bypassing intricate national security backgrounds and watermarks.
- •Date Normalization: Returns all dates (Birth, Issue, Expiry) in a standardized YYYY-MM-DD format, ready for JavaScript `Date` objects.
Sample JSON Output
The Node.js SDK returns an object matching this normalized JSON structure.
{
"success": true,
"data": {
"type": "passport",
"country_code": "BGR",
"nationality": "BGR",
"passport_number": "A1234567",
"surname": "IVANOV",
"given_names": "GEORGI",
"sex": "M",
"date_of_birth": "1993-06-18",
"place_of_birth": "SOFIA",
"date_of_issue": "2023-02-14",
"date_of_expiry": "2028-02-13",
"issuing_authority": "MINISTRY OF INTERIOR"
}
}Frequently Asked Questions
How does StructOCR compare to AWS Textract or Google Vision for Bulgarian documents?
Generic OCR services often struggle with the Cyrillic/English bilingual layout and the unique characters (ъ, ь, ю, я, щ) of Bulgarian passports, frequently misreading names or mixing up fields. Furthermore, you remain responsible for writing JavaScript parsing logic and validating MRZ checksums. StructOCR is a specialized API trained specifically on these Southeastern European documents, returning validated, labeled fields directly.
Do you store the uploaded images?
We do not store customer images. All data is processed in-memory (RAM) and is purged immediately after the API request is completed. We are a SOC2 compliant provider.
Can I pass a buffer instead of a file path in Node.js?
Yes. On Node.js server runtimes, the SDK accepts Buffer or Uint8Array, validates the decoded JPG, PNG, WebP, or PDF, and converts it to Base64 JSON before calling the REST API.
You May Also Like
Related tutorials, platform guides, and comparisons
Tunisia Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Tunisia passports.
Portugal Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Portugal passports.
Estonia Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Estonia passports.
Pakistan Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Pakistan passports.
Nigeria Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Nigeria passports.
Tajikistan Passport OCR with Node.js SDK
Use the official StructOCR Node.js SDK with TypeScript and Express to extract structured MRZ and VIZ data from Tajikistan passports.
Java Passport OCR API
Achieve 99%+ accuracy for Passport MRZ and VIZ data extraction in Java. Upload an image for a free test! Our API provides structured JSON output, handling glare and blur automatically.
C# Passport OCR API
Upload an image for a free test! High-accuracy C# Passport OCR API for parsing MRZ data. Get structured JSON output from passport images. Eliminate Tesseract errors and manual entry.
PHP Passport & Visa OCR API
Enterprise-grade PHP Passport & Visa OCR API. Test our instant JSON extraction for free. Accurate ICAO 9303 & identity document parsing with zero RegEx required.
Go Passport OCR API
Upload an image for a free test! Extract passport data via a simple REST API in Go. Get high-accuracy, ICAO-compliant JSON output from noisy images. Eliminate manual entry errors.
Technical Comparisons & Integrations
Explore platform integrations and competitive analysis
Heavy KYC SDKs vs. StructOCR Passport API for Mobile Apps
Discover why modern app developers are abandoning bulky on-device OCR SDKs in favor of lightweight, edge-deployed Passport APIs.
Add Passport OCR to Your Lovable.dev App in 5 Minutes
Step-by-step guide to integrating the StructOCR Passport scanner API into a Lovable.dev app through a secure server-side function for travel and KYC flows.
Add Passport OCR to Your Replit App in 5 Minutes
Step-by-step guide to integrating the StructOCR Passport scanner API into a Replit application. Build travel, hospitality, and global KYC flows using Replit Agent and Secrets.
Tesseract & PassportEye vs. StructOCR Passport API
A technical evaluation for developers deciding between open-source MRZ scanners and a dedicated cloud-based Passport OCR API.
Add Passport OCR to Your v0 App in 5 Minutes
Step-by-step guide to integrating the StructOCR Passport scanner API into a v0 by Vercel app through a secure Next.js server route.
AWS Textract vs. StructOCR API for Indian Passport Back Page Extraction
Why generic OCR fails at parsing the back page of Indian passports, and how a dedicated KYC API instantly extracts structured addresses, PIN codes, and family details.
From tutorial to production in 5 minutes.
You've seen the code. Now get your API key, grab your 200 free credits, and see it work with your own images. No credit card required.