choose as a deep-learning OCR model alternative with strong layout, table, and handwriting support.
OCR model that handles complex tables, forms, handwriting with full layout.
- 12.1k
- Python
- Apache-2.0
document OCR / visual text extraction
A deep-learning OCR model that converts document images to text by compressing visual context through a vision-language architecture.
Same problem, same approach. Swapping one for another is a config change, not a rewrite.
choose as a deep-learning OCR model alternative with strong layout, table, and handwriting support.
OCR model that handles complex tables, forms, handwriting with full layout.
choose as a deep-learning OCR model alternative with similar speed and accuracy goals.
GLM-OCR: Accurate × Fast × Comprehensive
Solves the same problem with a different architecture or at a different layer. Expect to rewrite the integration.
choose when you need a lightweight, mature OCR engine in C++ with no GPU dependency.
Tesseract Open Source OCR Engine (main repository)
choose when you need OCR directly in the browser or Node.js via Tesseract.
Pure Javascript OCR for more than 100 Languages 📖🎉🖥
choose when you want OCR through managed vision APIs rather than hosting a local model.
OCR & Document Extraction using vision models
choose when you need to call Tesseract OCR from Go.
Go package for OCR (Optical Character Recognition), by using Tesseract C++ library
choose when you need Tesseract OCR inside a PHP application.
A wrapper to work with Tesseract OCR inside PHP.
choose when you want to improve existing Tesseract output with LLM error correction and markdown formatting.
Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs
An earlier or more limited way to do the job. Still the right call when you need something small, proven, or CPU-only.
choose for Japanese manga text recognition specifically, not general document OCR.
Optical character recognition for Japanese text, with the main focus being Japanese manga
choose as a research/previous-generation deep-learning text recognition baseline.
Text recognition (optical character recognition) with deep learning methods, ICCV 2019
choose for academic document OCR to markdown with an older specialized model.
Implementation of Nougat Neural Optical Understanding for Academic Documents
Does the same job, but ships as an app. Useful to a person, not swappable into a codebase.
choose for a ready-made offline OCR GUI with batch and PDF support.
OCR software, free and offline. 开源、免费的离线OCR软件。支持截屏/批量导入图片,PDF文档识别,排除水印/页眉页脚,扫描/生成二维码。内置多国语言库。
choose for a cross-platform desktop GUI that combines translation and screen OCR.
🌈一个跨平台的划词翻译和OCR软件 | A cross-platform software for text translation and recognition.
choose for a macOS-native GUI with integrated OCR and translation.
Bob 是一款 macOS 平台的翻译和 OCR 软件。
choose for a ready-to-go Windows/WPF OCR and translation utility.
A ready-to-go translation ocr tool developed with WPF/WPF 开发的一款即用即走的翻译、OCR工具
choose for an all-in-one screenshot, offline OCR, and search desktop tool.
截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling screenshot Screen translator 支持Windows Linux macOS
choose for a desktop utility that combines screenshots, OCR, AI, and translation.
🚀 Screenshots, word marking, OCR, AI, translation software || 截图、划词、文字识别、AI、翻译软件
choose for a CLI tool that adds a searchable OCR text layer to scanned PDFs.
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
choose for an old experimental Android OCR app, not for code integration.
Experimental optical character recognition app
These projects were analysed and named DeepSeek-OCR among their alternatives. The relationship is not symmetric — how DeepSeek-OCR rates them is a separate judgement, made when DeepSeek-OCR is analysed in its own right.
calls DeepSeek-OCR “vision-model OCR library; choose for contextual compression and AI-native workflows”
PandaOCR - 多功能OCR图文识别+翻译+朗读+弹窗+公式+表格+图床+搜图+二维码
calls DeepSeek-OCR “Alternative OCR model with optical compression, for those needing a different OCR backend.”
Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs