when you want a lightweight cross-platform desktop OCR GUI (Java) with similar offline functionality.
树洞 OCR 文字识别(一款跨平台的 OCR 小工具)
- 3.1k
- Java
- LGPL-3.0
desktop offline OCR
Umi-OCR is a free, open-source offline desktop OCR application supporting screenshot recognition, batch image recognition, PDF recognition, QR codes, and formula recognition.
Same problem, same approach. Swapping one for another is a config change, not a rewrite.
when you want a lightweight cross-platform desktop OCR GUI (Java) with similar offline functionality.
树洞 OCR 文字识别(一款跨平台的 OCR 小工具)
when you prefer a minimal, engine-specific GUI for PaddleOCR rather than Umi-OCR's feature set.
The GUI for "paddlepaddle" OCR
when you want a cross-platform screenshot tool that also provides offline OCR, search, and translation in one app.
截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling screenshot Screen translator 支持Windows Linux macOS
when you need a classic Tianruo OCR-style desktop app with local PaddleOCR recognition.
天若ocr开源版本的本地版,采用Chinese-lite和paddleocr识别框架
Solves the same problem with a different architecture or at a different layer. Expect to rewrite the integration.
when you need a self-hosted PaddleOCR HTTP server instead of a desktop GUI.
基于PaddleOCR搭建的OCR server... 离线部署用
when you need OCR within an iOS/Swift app rather than a desktop tool.
PaddleOCR是一款应用于iOS设备上的通用文字识别的OCR库.
when you need a fast Python library with a small ONNX model (20MB) to embed in your own app.
Python3 package for Chinese/English OCR,use paddleocr-v5 onnx model(~20MB), with ultra-fast inference speed. 基于ppocr-v5-onnx模型推理,中英文OCR开源SOTA,推理速度超快。
when you want to use Chromium's screen-ai OCR via a Python wrapper instead of a full app.
Wrapper for Chromium screen-ai OCR
when you need OCR functionality in a Go service or application using ONNX.
go-ocr 是一款基于 Golang + ONNX 构建的 OCR 工具库,专注于为 Go 生态提供简单易用、可扩展的文字识别能力。
when you want a unified API wrapper to switch between multiple OCR engines without changing your code.
MyLittleOCR 是一个统一的 OCR 库包装器,提供一致的 API,便于集成和切换多个 OCR 引擎。 MyLittleOCR is a unified OCR wrapper providing a consistent API for seamless integration and switching between multiple OCR engines.
when you need a simple, self-hosted web UI/API for offline OCR instead of a desktop app.
开源的中英文离线 OCR,使用 PaddleOCR 实现,提供了简单的 Web 页面及接口
when you need OCR as a .NET local library and may require table recognition.
PaddleOCRSharp是一个.NET的OCR工具本地类库,可离线使用。包含文本识别、文本检测、表格识别功能。本项目针对小图识别不准的情况下做了优化,比飞桨原代码识别准确率有所提高。 包含总模型仅8.6M的超轻量级中文OCR,单模型支持中英文数字组合识别、竖排文本识别、长文本识别。同时支持多种文本检测。
when you need offline OCR in a Node.js/JavaScript project via a PaddleOCR-based library.
基于paddleOCR的nodejs库
when you want OCR deployed as a serverless API on the cloud rather than running a local desktop app.
基于 Serverless 架构部署通用文字识别 PaddleOCR
when you need to serve any OCR model as an online API endpoint rather than using a packaged app.
Turn any OCR models into online inference API endpoint 🚀 🌖
when you want a simple Flask API wrapper around PaddleOCR for integration with other services.
✅Deploy PaddleOCR with flask | 利用Flask对PaddleOCR进行部署,方便调用
An earlier or more limited way to do the job. Still the right call when you need something small, proven, or CPU-only.
for a lightweight Korean-only OCR script if you don't need the full desktop app.
This is a Korean OCR Python code using the paddleOCR library
for a quick screenshot-only OCR workflow integrated with Snipaste, simpler than Umi-OCR.
基于Snipaste的截图文字识别工具,使用飞浆的OCR模型,基于Snipaste的强大截图功能,实现截图文字自动识别。
Does the same job, but ships as an app. Useful to a person, not swappable into a codebase.
when you specifically need invoice OCR output to Excel, not general-purpose OCR.
基于OCR技术的自动识别发票内容,导出到Excel。(自动识别图片、pdf文件)
for converting scanned Chinese ancient books to searchable PDFs with specialized layout handling.
智能 OCR 工具 - 将扫描版 PDF 转换为可全文搜索的 PDF,专为中文古籍、学术文献设计
for a web-based tool focused on PDF text extraction with OCR, not a general-purpose desktop app.
PDF text data extraction web app with OCR for scanned documents
for fast hard-subtitle extraction from video on Apple silicon/NVIDIA GPUs, a niche OCR workflow.
快如闪电的硬字幕提取工具。仅需苹果M1芯片或英伟达3060显卡即可达到10倍速提取。A very fast tool for video hardcode subtitle extraction
These projects were analysed and named Umi-OCR among their alternatives. The relationship is not symmetric — how Umi-OCR rates them is a separate judgement, made when Umi-OCR is analysed in its own right.
calls Umi-OCR “offline OCR desktop app; choose for privacy/offline use and modern UI”
PandaOCR - 多功能OCR图文识别+翻译+朗读+弹窗+公式+表格+图床+搜图+二维码
calls Umi-OCR “Direct offline OCR desktop app with screenshot, batch, and PDF support, similar feature set to the seed.”
树洞 OCR 文字识别(一款跨平台的 OCR 小工具)
calls Umi-OCR “OCR-only tool; choose if you need pure offline OCR without translation.”
🌈一个跨平台的划词翻译和OCR软件 | A cross-platform software for text translation and recognition.
calls Umi-OCR “Choose it for an offline OCR-focused tool with batch and PDF support, without translation.”
Bob 是一款 macOS 平台的翻译和 OCR 软件。
calls Umi-OCR “Feature-rich OCR-only offline tool, narrower than STranslate.”
A ready-to-go translation ocr tool developed with WPF/WPF 开发的一款即用即走的翻译、OCR工具
calls Umi-OCR “a desktop OCR GUI app for end users, not a library for developers.”
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
calls Umi-OCR “choose for a ready-made offline OCR GUI with batch and PDF support.”
Contexts Optical Compression
calls Umi-OCR “offline GUI OCR application, not a model you can import”
Official code implementation of General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model
calls Umi-OCR “End-user desktop OCR app, not embeddable in a codebase.”
Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs