opendatalab/MinerUπ₯ active
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
Open-source platform for extracting structured data from documents using AI.
$ git clone https://github.com/DocumindHQ/documind.gitTransforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
[ECCV 2026] A diffusion-based framework for document OCR that replaces autoregressive decoding with block-level parallel diffusion decoding.
The official repo for βDolphin: Document Image Parsing via Heterogeneous Anchor Promptingβ, ACL, 2025.
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
Free and open-source PDF editor for Windows with a built-in PDF 2.0 engine. View, annotate, OCR, merge, split, crop, rotate, compare, edit text, draw, sign, fill forms, print, flatten, and open password-protected PDFs without a subscription.
A curated collection of practical AI projects implementing OCR systems, RAG, AI agents, and other AI use cases.
Data from GitHub Β· snapshot Sep 24, 2026