Key-value extraction
Capture named fields and their values from varied document layouts.
Classify documents, extract fields and tables, preserve layout context, and return application-ready JSON—built for Southeast Asian formats, scripts, and workflows.
{
"document_type": "invoice",
"language": "vi",
"fields": {
"supplier": "...",
"invoice_no": "...",
"tax": "...",
"total": "..."
},
"line_items": [...]
}One document pipeline
TurboLens combines regional OCR, layout understanding, and configurable extraction so teams can move from unstructured files to usable application data.
Send PDFs, scans, or document images through an API-first processing flow.
Identify the document type and route it to the appropriate extraction path.
Capture configured fields, tables, layout structure, stamps, and visual elements.
Return structured JSON shaped for applications, data stores, and reviewer tools.
Beyond basic OCR
Basic OCR returns characters. Document Intelligence preserves the relationships between fields, tables, layout regions, and visual signals so the output is easier to use in real workflows.
Capture named fields and their values from varied document layouts.
Preserve row, column, and header relationships for downstream processing.
Recognize sections, forms, checkboxes, figures, and multi-column structures.
Locate official markings and include their regions in reviewer workflows.
Reduce watermark and background interference before text extraction.
Turn relevant visual elements into structured, reviewable information.
Regional document coverage
Supplier details, tax fields, line items, totals, and payment references.
Explore documentsRegional national IDs, passports, bilingual fields, and machine-readable zones.
Explore documentsParties, clauses, dates, obligations, signatures, and structured sections.
Explore documentsBills of lading, packing lists, customs declarations, and delivery records.
Explore documentsTax forms, applications, medical certificates, and government documents.
Explore documentsApply it to your workflow
Explore how extraction, classification, and structured outputs support insurance, logistics, public-sector, and custom enterprise workflows.
Document Intelligence combines OCR with document classification, layout understanding, and structured extraction. It turns PDFs, scans, and document images into organized data that applications and reviewers can use.
Basic OCR primarily returns text. TurboLens also captures field relationships, tables, document sections, and configured output structures so the result fits downstream workflows.
TurboLens supports regional invoices, receipts, identity documents, passports, contracts, tax forms, customs documents, medical certificates, and other configured enterprise document types.
Teams submit documents through API-first workflows and receive structured JSON. Output schemas and processing logic can be aligned to the fields and document variants used by the application.