🏆 #467 overall#77 of 313 in OCR & Documents

alisen39 /TrWebOCR

开源易用的中文离线OCR,识别率媲美大厂,并且提供了易用的web页面及web的接口,方便人类日常工作使用或者其他程序来调用~

$ git clone https://github.com/alisen39/TrWebOCR.git
GitHub social preview for alisen39/TrWebOCR
Stars
2.9K
2,880
Forks
622
622
Language
Python
License
Apache-2.0
Created
May 1, 2020
6.4 years old
Last push
Jun 14, 2023

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

ianzhao/textshot

Python tool for grabbing text via screenshot

PythonMITupdated Dec 20, 2024
GitHub ↗★ 1.8K⑂ 255
OCR & Documents

yinchangchang/ocr_densenet

第一届西安交通大学人工智能实践大赛(2018AI实践大赛--图片文字识别)第一名;仅采用densenet识别图中文字

Pythonno licenseupdated Mar 12, 2019
GitHub ↗★ 466⑂ 160
OCR & Documents

ocrmypdf/OCRmyPDF🔥 active

OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched

PythonMPL-2.0updated Sep 22, 2026
GitHub ↗★ 34.9K⑂ 2.4K
OCR & Documents

run-llama/liteparse🔥 active

A fast, helpful, and open-source document parser

RustApache-2.0updated Sep 22, 2026
GitHub ↗★ 12.6K⑂ 860
OCR & Documents

pymupdf/PyMuPDF🔥 active

PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.

PythonAGPL-3.0updated Sep 24, 2026
GitHub ↗★ 10.8K⑂ 804
OCR & Documents

bytedance/Dolphin

The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.

PythonOtherupdated Mar 25, 2026
GitHub ↗★ 9.1K⑂ 778

Data from GitHub · snapshot Sep 24, 2026