All articlesTechnology

How AI OCR and Document Translation Actually Work

July 19, 2026 5 min read

Abstract visualisation of an AI scan line reading a document

Document translation is not the same as translating a paragraph of text. The hard part is reading a scan reliably and preserving the structure that makes an official document meaningful.

Step 1 — Reading the page

A multimodal model looks at the page image directly rather than relying on a separate OCR pass, which helps with stamps, low-contrast scans, mixed scripts and handwriting. The output is text plus positional context: headings, tables, field labels and values.

Step 2 — Understanding the document type

A birth certificate and a court order have different conventions. Detecting the document type lets the system apply the right terminology and keep field labels consistent throughout a multi-page file.

Step 3 — Translating with constraints

  • Names and identifiers are preserved, not translated.
  • Stamps and seals are described in brackets.
  • Ambiguous or unreadable content is flagged rather than guessed.
  • Layout markers are kept so the output can be shown side by side.

Step 4 — Confidence scoring

Each translation gets a confidence score based on scan quality, ambiguity and unresolved segments. A low score is a signal to order human review — not a claim that the text is wrong. A high score is not a substitute for certification when an institution requires it.

Translate this document today

AI translation from $1.29 per page. Certified human review from $24.00 per page. Prices shown in USD for your region; suggested target language: English.

Acceptance requirements vary by institution. Always confirm the format your receiving office requires.