🏆 #1,618 overall#248 of 313 in OCR & Documents

vorojar /Folio-OCR

Open-source batch OCR workbench — a free, local alternative to ABBYY FineReader. Powered by Ollama + GLM-OCR + PP-DocLayoutV3, ~0.5s/page on RTX 4090. Three-panel editor, layout-aware, PDF/image batch processing, Markdown/Word export. 批量OCR工作台,纯本地运行,免费平替ABBYY,适合书籍文档数字化。

$ git clone https://github.com/vorojar/Folio-OCR.git
GitHub social preview for vorojar/Folio-OCR
Stars
474
474
Forks
62
62
Language
Python
License
MIT
Created
Feb 7, 2026
0.6 years old
Last push
Jun 18, 2026

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

run-llama/liteparse🔥 active

A fast, helpful, and open-source document parser

RustApache-2.0updated Sep 22, 2026
GitHub ↗★ 12.6K⑂ 860
OCR & Documents

oomol-lab/pdf-craft🔥 active

PDF craft can convert PDF files into various other formats. This project will focus on processing PDF files of scanned books.

PythonMITupdated Sep 23, 2026
GitHub ↗★ 6.3K⑂ 458
OCR & Documents

PaddlePaddle/PaddleX

All-in-One Development Tool based on PaddlePaddle

🎙️ Speech Recognition
PythonApache-2.0updated Jun 25, 2026
GitHub ↗★ 6.3K⑂ 1.2K
OCR & Documents

TheJoeFin/Text-Grab🔥 active

Use OCR in Windows quickly and easily with Text Grab. With optional background process and notifications.

C#MITupdated Sep 22, 2026
GitHub ↗★ 5K⑂ 331
OCR & Documents

wxyhgk/retain-pdf🔥 active

在保留版面、公式与结构的前提下进行 PDF 翻译,适用于科研与技术文档

PythonMITupdated Sep 21, 2026
GitHub ↗★ 2.3K⑂ 278

Data from GitHub · snapshot Sep 24, 2026