Skip to content
Smart Document & Digital Productivity Platform

OCR

Read the text inside images and PDFs and turn it into editable, searchable content.

Drop your documents here

or

PDF • JPG • PNG • WEBP • TIFF • BMP

Up to 25 MB per file · 10 files at once

No account needed. Your files stay private.

Read the text inside images and PDFs and turn it into editable, searchable content.

What it does

Zegi's OCR extracts text from images and PDFs and, where it recognizes them, pulls structured fields from invoices, receipts and passports — returning real confidence values, never fabricated data.

How it works

  1. Upload an image or PDF.
  2. Zegi reads it with PaddleOCR (primary) and Tesseract (fallback), detecting language automatically.
  3. Review the text and structured fields, then copy or export to Word, Excel or PDF.

Supported formats

Input: images (PNG, JPG/JPEG, WebP) and PDFs. Output: text, structured data, or exports.

What you get

Editable text plus any detected structured data, ready to copy or export.

Important limitations

  • Accuracy depends on the source quality; low-resolution or noisy scans reduce results.
  • Uncertain fields are returned as empty rather than guessed — verify important values.

Privacy & your files

Your files are processed on Zegi's own servers — never sent to a third-party AI service — and are removed automatically after processing. See our Privacy Policy and Data Retention Policy for details.

Frequently asked questions

Is OCR free?
Public OCR is free to use; accounts and the API unlock higher limits and additional features. See Pricing.
Which languages are supported?
English and Arabic, including mixed-language documents.
What about invoices and passports?
Where these are recognized, Zegi extracts structured fields; anything it can't read confidently is left empty rather than guessed.
Is there an API?
Yes — POST /api/v1/ocr, with async jobs and webhooks. See the API documentation.

Related tools

Helpful links