Feature

Multilingual OCR

Read documents in many languages and scripts, including Arabic.

What is Multilingual OCR in Mahad OCR?

Available now. The AI reader is not restricted to Latin script. Input: Documents in supported languages as PNG, JPG or PDF. Limitation: We publish no per-language accuracy table.

Available now

What it does

The AI reader is not restricted to Latin script. Arabic documents have been processed successfully in production, and field labels are returned in normalized English keys regardless of the source language.

How it works

  1. 1Upload the document — no language parameter needed
  2. 2Language and script are handled by the reader
  3. 3Values are returned as printed; keys are normalized

Sample result

Illustrative response shape using synthetic data.

{ "document_type": "certificate", "values": { "full_name": "…" }, "field_evidence": { "full_name": "corroborated" } }

Supported inputs

Documents in supported languages as PNG, JPG or PDF.

Fields returned

  • same normalized keys regardless of source language
  • original values preserved

API

Automatic — no language flag required.

Security & privacy

Identical processing and isolation rules for every language.

Current limitations

We publish no per-language accuracy table. Arabic has been verified in production; other scripts should be validated on your own documents before you depend on them. The free on-server engine is Latin-oriented — non-Latin documents use the AI reader.

Try it on your own document

Run the live demo, or create an account and get an API key.