OCR: Documents
Collection
Roughly ordered by recent releases, useful updates and current usage. Practical OCR and document parsing models. Reviewed September 2026. • 17 items • Updated • 1
Curated OCR models for documents, languages, handwriting and text in images. Browse four collections with short practical notes.
Note Scanned pages, books, forms, tables and formulas. Roughly ordered by recent releases, useful updates and current usage; some models need a full parsing pipeline.
Note Language-specific OCR, including Thai, Japanese, Vietnamese, Arabic, Korean and Devanagari. Check each model's expected input and domain.
Note Handwriting and historical print, including Swedish, Norwegian, German Kurrent, Tibetan and Hebrew-script manuscripts. Many models need cropped lines.
Note Text-line and region recognition, including Kraken and PaddleOCR, plus pipelines that find and read text in documents and photographs.