oomol-lab/pdf-craft

PDF craft can convert PDF files into various other formats. This project will focus on processing PDF files of scanned books.

View on GitHub ↗Jump to charts ↓Open shareable report

Summary Information

Updated 25 minutes ago
Added to GitGenius on September 10th, 2026
Created on February 12th, 2025
Open Issues & Pull Requests: 52 (+0)
GitHub issues: Enabled
Number of forks: 450
Total Stargazers: 6,290 (+0)
Total Subscribers: 21 (+0)

Repository Insights (GitGenius)

Median issue/PR response: 5.0 hours
Mean response time: 17.5 hours
90th percentile: 33.0 hours
Tracked items: 110

How this project is maintained

Around half of the issues opened in the past year never receive a reply. 100% of open issues come from outside the core team, so the backlog reflects real-world use rather than internal planning. Only 9% of issues opened in the past year have been closed. Three people close 91% of everything that gets resolved.

Charts & Analytics

Fetching additional details & charts...

Issue Activity (beta)

Open issues: 47
New in 7 days: 1
Closed in 7 days: 0
Avg open age: 367 days
Stale 30+ days: 46
Stale 90+ days: 43

Recent activity

Opened in 7 days: 1
Closed in 7 days: 0
Comments in 7 days: 1
Events in 7 days: 1

Top labels

  • wontfix (2)

Most active issues this week

Detailed Description

PDF Craft is a document processing tool that converts PDF files into various other formats with a focus on scanned book materials.

The tool addresses the challenge of extracting and transforming content from scanned PDF documents, which often require optical character recognition to become machine-readable and usable in other formats. PDF Craft handles this by integrating OCR capabilities to process scanned book PDFs and convert them into alternative formats suitable for different use cases.

Developers working with digitized book collections or archival scanned documents should consider PDF Craft if they need to extract text and convert PDFs into formats beyond their original state. The project is particularly suited for workflows involving large volumes of scanned materials that require format conversion and text extraction. The tool's focus on scanned books rather than general PDF processing means it is optimized for the specific challenges of document images with variable quality and layout complexity.

The project shows active development with regular commits and ongoing refinement of its core functionality. The codebase receives consistent updates addressing both feature additions and maintenance needs. The project maintains an associated homepage that provides additional documentation and resources for users exploring the tool's capabilities.