The full upstream README, mirrored here for reference. Install config, tool schemas, adoption signals, and an original overview live on the Ocr listing page.
Extract text from images and PDFs (best-in-class Arabic, manga-aware Japanese, 13+ languages auto-detected), translate 100+ languages, pull structured fields from invoices/receipts/IDs — with Saudi ZATCA e-invoice QR validation — and get RAG-ready markdown or semantic chunks. Self-hosted, PDPL-compliant: documents are processed transiently and never stored.
No signup needed. Your first call auto-provisions a free trial key.
Registry: com.auto-reader/ocr · Tools: ocr_image, extract_document,
translate_text, ocr_and_translate, get_usage, create_api_key.
Idempotency-Key header honored; 402 responses carry a buy_direct link.format=markdown / format=chunks (title-bounded, tables kept whole) turn
any scan into RAG input in one call; Accept: text/markdown returns a raw
.md body.extract_document returns {value, confidence, box} per field, sourced
from OCR geometry — never model guesswork.Site: https://ocr.auto-reader.com · العربية: https://ocr.auto-reader.com/ar