OCR & Documents
๐ #488 overall#81 of 313 in OCR & Documents๐ฅ active this week
dynobo /normcap
OCR powered screen-capture tool to capture information instead of images
$ git clone https://github.com/dynobo/normcap.gitStars
2.7K
2,730
Forks
125
125
Language
Python
License
Other
Created
Aug 14, 2019
7.1 years old
Last push
Sep 18, 2026
๐ฅ this week
Categories
GitHub topics
More in OCR & Documents
OCR & Documents
hiroi-sora/Umi-OCR
OCR software, free and offline. ๅผๆบใๅ ่ดน็็ฆป็บฟOCR่ฝฏไปถใๆฏๆๆชๅฑ/ๆน้ๅฏผๅ ฅๅพ็๏ผPDFๆๆกฃ่ฏๅซ๏ผๆ้คๆฐดๅฐ/้กต็้กต่๏ผๆซๆ/็ๆไบ็ปด็ ใๅ ็ฝฎๅคๅฝ่ฏญ่จๅบใ
PythonMITupdated Nov 20, 2025
OCR & Documents
ocrmypdf/OCRmyPDF๐ฅ active
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
PythonMPL-2.0updated Sep 22, 2026
OCR & Documents
pymupdf/PyMuPDF๐ฅ active
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
PythonAGPL-3.0updated Sep 24, 2026
OCR & Documents
bytedance/Dolphin
The official repo for โDolphin: Document Image Parsing via Heterogeneous Anchor Promptingโ, ACL, 2025.
PythonOtherupdated Mar 25, 2026
OCR & Documents
axa-group/Parsr
Transforms PDF, Documents and Images into Enriched Structured Data
๐ง Natural Language Processing
JavaScriptApache-2.0updated Mar 20, 2026
Data from GitHub ยท snapshot Sep 24, 2026