hiroi-sora/umi-ocr

OCR software, free and offline. 开源、免费的离线OCR软件。支持截屏/批量导入图片,PDF文档识别,排除水印/页眉页脚,扫描/生成二维码。内置多国语言库。

View on GitHub ↗Jump to charts ↓Open shareable report

Summary Information

Updated 13 minutes ago
Added to GitGenius on August 31st, 2026
Created on March 28th, 2022
Open Issues & Pull Requests: 358 (+0)
Number of forks: 4,599
Total Stargazers: 47,011 (+0)
Total Subscribers: 227 (+0)

Repository Insights (GitGenius)

Median issue/PR response: 15.8 hours
Mean response time: 9.7 days
90th percentile: 22.9 days
Tracked items: 282

How this project is maintained

Around half of the issues opened in the past year never receive a reply. 99% of open issues come from outside the core team, so the backlog reflects real-world use rather than internal planning. 75% of tracked open issues have had no activity in three months, so the open count overstates what is actively being worked. Only 3% of issues opened in the past year have been closed. Three people close 60% of everything that gets resolved.

Charts & Analytics

Fetching additional details & charts...

Issue Activity (beta)

Open issues: 289
New in 7 days: 0
Closed in 7 days: 0
Avg open age: 474 days
Stale 30+ days: 286
Stale 90+ days: 262

Recent activity

Opened in 7 days: 0
Closed in 7 days: 0
Comments in 7 days: 0
Events in 7 days: 0

Top labels

  • 功能建议 (6)
  • bug (4)
  • 模型效果 (3)
  • 系统适配 (1)

Most active issues this week

Detailed Description

Umi-OCR is an offline optical character recognition tool that runs locally without requiring network connectivity.

The tool addresses the need for free, accessible text extraction from images and documents. It uses PaddleOCR as its recognition engine and provides a graphical interface built with Qt and QML. The software handles multiple input methods including screenshot capture, batch image import, and PDF document processing. It includes built-in language libraries for multilingual recognition and offers text post-processing capabilities to handle different document layouts and exclude unwanted regions like watermarks, headers, and footers. The tool also supports QR code scanning and generation.

Umi-OCR suits developers and end users who need reliable offline OCR without cloud dependencies or subscription costs. It works well for batch processing workflows, document digitization, and integration into other applications through its command-line interface and HTTP API. The software targets Windows 7 x64 and Linux x64 systems. Unlike cloud-based OCR services, this tool operates entirely offline and requires no network access after installation.

The project maintains active development with regular updates and multilingual support through community translation efforts. Bug reports and feature requests receive attention through the project's issue tracker. The codebase is organized to support both end-user applications and developer integration, with documentation provided for command-line usage and HTTP interface implementation. The project includes build instructions for developers working with Windows and Linux environments.