OCR & Documents
opendatalab/MinerU🔥 active
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
PythonOtherupdated Sep 24, 2026
OCR & Document Extraction using vision models
$ git clone https://github.com/getomni-ai/zerox.gitTransforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
A fast, helpful, and open-source document parser
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.
飞鼠格式 FlyingMouse Format - Windows 免费文件格式转换工具(离线可用,内置 FFmpeg/LibreOffice/Poppler/Tesseract)。图片/文档/表格/PPT/PDF/音视频/WPS 格式互转 + OCR + 批量转换;音频仅支持普通格式。作者:牢蜂(LaoFeng)|仅供个人免费使用,禁止商业售卖/转卖/套壳
Data from GitHub · snapshot Sep 24, 2026