AltHub

hiroi-sora/Umi-OCR alternatives

desktop offline OCR

Umi-OCR is a free, open-source offline desktop OCR application supporting screenshot recognition, batch image recognition, PDF recognition, QR codes, and formula recognition.

46.7kPythonMITslowing · last push 9mo agoGitHub

Drop-in peers

Same problem, same approach. Swapping one for another is a config change, not a rewrite.

AnyListen/tools-ocr

when you want a lightweight cross-platform desktop OCR GUI (Java) with similar offline functionality.

树洞 OCR 文字识别(一款跨平台的 OCR 小工具)

3.1k
Java
LGPL-3.0
dormant2.3y ago
cloudy-sfu/GUI-for-paddlepaddle-OCR

when you prefer a minimal, engine-specific GUI for PaddleOCR rather than Umi-OCR's feature set.

The GUI for "paddlepaddle" OCR

57
Python
GPL-3.0
slowing10mo ago
xushengfeng/eSearch

when you want a cross-platform screenshot tool that also provides offline OCR, search, and translation in one app.

截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling screenshot Screen translator 支持Windows Linux macOS

7.0k
TypeScript
GPL-3.0
active5d ago
wangfreexx/wangfreexx-tianruoocr-cl-paddle

when you need a classic Tianruo OCR-style desktop app with local PaddleOCR recognition.

天若ocr开源版本的本地版,采用Chinese-lite和paddleocr识别框架

1.4k
C#
GPL-3.0
dormant2.2y ago

Same job, different approach

Solves the same problem with a different architecture or at a different layer. Expect to rewrite the integration.

PantsuDango/DangoOCR

when you need a self-hosted PaddleOCR HTTP server instead of a desktop GUI.

基于PaddleOCR搭建的OCR server... 离线部署用

261
Python
none
dormant3.4y ago
Leonard-iOS/PaddleOCR

when you need OCR within an iOS/Swift app rather than a desktop tool.

PaddleOCR是一款应用于iOS设备上的通用文字识别的OCR库.

38
Swift
MIT
dormant4.8y ago
shibing624/imgocr

when you need a fast Python library with a small ONNX model (20MB) to embed in your own app.

Python3 package for Chinese/English OCR,use paddleocr-v5 onnx model(~20MB), with ultra-fast inference speed. 基于ppocr-v5-onnx模型推理,中英文OCR开源SOTA,推理速度超快。

134
Python
Apache-2.0
active4mo ago
sergiocorreia/clv-locro

when you want to use Chromium's screen-ai OCR via a Python wrapper instead of a full app.

Wrapper for Chromium screen-ai OCR

42
Python
MIT
active3mo ago
GetcharZp/go-ocr

when you need OCR functionality in a Go service or application using ONNX.

go-ocr 是一款基于 Golang + ONNX 构建的 OCR 工具库,专注于为 Go 生态提供简单易用、可扩展的文字识别能力。

71
Go
MIT
active3d ago
X-T-E-R/my-little-ocr

when you want a unified API wrapper to switch between multiple OCR engines without changing your code.

MyLittleOCR 是一个统一的 OCR 库包装器,提供一致的 API,便于集成和切换多个 OCR 引擎。 MyLittleOCR is a unified OCR wrapper providing a consistent API for seamless integration and switching between multiple OCR engines.

54
Python
MIT
dormant1.9y ago
lewangdev/PaddleWebOCR

when you need a simple, self-hosted web UI/API for offline OCR instead of a desktop app.

开源的中英文离线 OCR,使用 PaddleOCR 实现,提供了简单的 Web 页面及接口

132
Vue
Apache-2.0
dormant4.3y ago
raoyutian/PaddleOCRSharp

when you need OCR as a .NET local library and may require table recognition.

PaddleOCRSharp是一个.NET的OCR工具本地类库,可离线使用。包含文本识别、文本检测、表格识别功能。本项目针对小图识别不准的情况下做了优化,比飞桨原代码识别准确率有所提高。 包含总模型仅8.6M的超轻量级中文OCR,单模型支持中英文数字组合识别、竖排文本识别、长文本识别。同时支持多种文本检测。

145
C#
none
active2mo ago
xushengfeng/eSearch-OCR

when you need offline OCR in a Node.js/JavaScript project via a PaddleOCR-based library.

基于paddleOCR的nodejs库

127
TypeScript
Apache-2.0
active10d ago
duolabmeng6/paddlehub_ppocr

when you want OCR deployed as a serverless API on the cloud rather than running a local desktop app.

基于 Serverless 架构部署通用文字识别 PaddleOCR

126
Python
Apache-2.0
slowing12mo ago
bentoml/BentoOCR

when you need to serve any OCR model as an online API endpoint rather than using a packaged app.

Turn any OCR models into online inference API endpoint 🚀 🌖

60
Python
none
active1mo ago
WangRongsheng/PaddleOCR-Flask-deploy

when you want a simple Flask API wrapper around PaddleOCR for integration with other services.

✅Deploy PaddleOCR with flask | 利用Flask对PaddleOCR进行部署,方便调用

44
HTML
none
dormant4.2y ago

Older generation, or narrower

An earlier or more limited way to do the job. Still the right call when you need something small, proven, or CPU-only.

yunwoong7/korean_ocr_using_paddleOCR

for a lightweight Korean-only OCR script if you don't need the full desktop app.

This is a Korean OCR Python code using the paddleOCR library

26
Jupyter Notebook
Apache-2.0
dormant3.1y ago
formero009/SnipasteOCR

for a quick screenshot-only OCR workflow integrated with Snipaste, simpler than Umi-OCR.

基于Snipaste的截图文字识别工具,使用飞浆的OCR模型,基于Snipaste的强大截图功能,实现截图文字自动识别。

126
Python
MIT
slowing1.3y ago

End-user tools

Does the same job, but ships as an app. Useful to a person, not swappable into a codebase.

zhiweiiii/fapiao-ocr-excel

when you specifically need invoice OCR output to Excel, not general-purpose OCR.

基于OCR技术的自动识别发票内容,导出到Excel。(自动识别图片、pdf文件)

35
Python
none
slowing10mo ago
anon-research-tools/intelligent-ocr

for converting scanned Chinese ancient books to searchable PDFs with specialized layout handling.

智能 OCR 工具 - 将扫描版 PDF 转换为可全文搜索的 PDF,专为中文古籍、学术文献设计

30
Python
none
active6mo ago
nainiayoub/pdf-text-data-extractor

for a web-based tool focused on PDF text extraction with OCR, not a general-purpose desktop app.

PDF text data extraction web app with OCR for scanned documents

96
Python
none
dormant2.2y ago
nhjydywd/SubtitleOCR

for fast hard-subtitle extraction from video on Apple silicon/NVIDIA GPUs, a niche OCR workflow.

快如闪电的硬字幕提取工具。仅需苹果M1芯片或英伟达3060显卡即可达到10倍速提取。A very fast tool for video hardcode subtitle extraction

688
Swift
GPL-3.0
slowing8mo ago

Listed as an alternative to

These projects were analysed and named Umi-OCR among their alternatives. The relationship is not symmetric — how Umi-OCR rates them is a separate judgement, made when Umi-OCR is analysed in its own right.

miaomiaosoft/PandaOCR

calls Umi-OCRoffline OCR desktop app; choose for privacy/offline use and modern UI

PandaOCR - 多功能OCR图文识别+翻译+朗读+弹窗+公式+表格+图床+搜图+二维码

5.3kdormant
AnyListen/tools-ocr

calls Umi-OCRDirect offline OCR desktop app with screenshot, batch, and PDF support, similar feature set to the seed.

树洞 OCR 文字识别(一款跨平台的 OCR 小工具)

3.1kJavadormant
pot-app/pot-desktop

calls Umi-OCROCR-only tool; choose if you need pure offline OCR without translation.

🌈一个跨平台的划词翻译和OCR软件 | A cross-platform software for text translation and recognition.

19.3kJavaScriptactive
ripperhe/Bob

calls Umi-OCRChoose it for an offline OCR-focused tool with batch and PDF support, without translation.

Bob 是一款 macOS 平台的翻译和 OCR 软件。

9.7kslowing
STranslate/STranslate

calls Umi-OCRFeature-rich OCR-only offline tool, narrower than STranslate.

A ready-to-go translation ocr tool developed with WPF/WPF 开发的一款即用即走的翻译、OCR工具

7.8kC#active
PaddlePaddle/PaddleOCR

calls Umi-OCRa desktop OCR GUI app for end users, not a library for developers.

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

88.1kPythonactive
deepseek-ai/DeepSeek-OCR

calls Umi-OCRchoose for a ready-made offline OCR GUI with batch and PDF support.

Contexts Optical Compression

23.8kPythonslowing
Ucas-HaoranWei/GOT-OCR2.0

calls Umi-OCRoffline GUI OCR application, not a model you can import

Official code implementation of General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model

8.2kPythondormant
Dicklesworthstone/llm_aided_ocr

calls Umi-OCREnd-user desktop OCR app, not embeddable in a codebase.

Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown formatting of scanned PDFs

3.0kPythonactive