🏆 #127 overall#20 of 313 in OCR & Documents

getomni-ai /zerox

OCR & Document Extraction using vision models

$ git clone https://github.com/getomni-ai/zerox.git
GitHub social preview for getomni-ai/zerox
Stars
12.3K
12,263
Forks
848
848
Language
TypeScript
License
MIT
Created
Jul 21, 2024
2.2 years old
Last push
May 20, 2025

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

opendatalab/MinerU🔥 active

Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.

PythonOtherupdated Sep 24, 2026
GitHub ↗★ 80.6K⑂ 6.7K
OCR & Documents

ocrmypdf/OCRmyPDF🔥 active

OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched

PythonMPL-2.0updated Sep 22, 2026
GitHub ↗★ 34.9K⑂ 2.4K
OCR & Documents

run-llama/liteparse🔥 active

A fast, helpful, and open-source document parser

RustApache-2.0updated Sep 22, 2026
GitHub ↗★ 12.6K⑂ 860
OCR & Documents

pymupdf/PyMuPDF🔥 active

PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.

PythonAGPL-3.0updated Sep 24, 2026
GitHub ↗★ 10.8K⑂ 804
OCR & Documents

bytedance/Dolphin

The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.

PythonOtherupdated Mar 25, 2026
GitHub ↗★ 9.1K⑂ 778
OCR & Documents

LaoFeng-mouse/flyingmouse-format🔥 active

飞鼠格式 FlyingMouse Format - Windows 免费文件格式转换工具(离线可用,内置 FFmpeg/LibreOffice/Poppler/Tesseract)。图片/文档/表格/PPT/PDF/音视频/WPS 格式互转 + OCR + 批量转换;音频仅支持普通格式。作者:牢蜂(LaoFeng)|仅供个人免费使用,禁止商业售卖/转卖/套壳

JavaScriptOtherupdated Sep 21, 2026
GitHub ↗★ 6.2K⑂ 517

Data from GitHub · snapshot Sep 24, 2026