πŸ† #1,405 overall#211 of 313 in OCR & Documents

opendatalab /MinerU-Diffusion

[ECCV 2026] A diffusion-based framework for document OCR that replaces autoregressive decoding with block-level parallel diffusion decoding.

$ git clone https://github.com/opendatalab/MinerU-Diffusion.git
GitHub social preview for opendatalab/MinerU-Diffusion
Stars
619
619
Forks
42
42
Language
Python
License
MIT
Created
Mar 13, 2026
0.5 years old
Last push
Jun 18, 2026

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

opendatalab/MinerUπŸ”₯ active

Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.

PythonOtherupdated Sep 24, 2026
GitHub β†—β˜… 80.6Kβ‘‚ 6.7K
OCR & Documents

bytedance/Dolphin

The official repo for β€œDolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.

PythonOtherupdated Mar 25, 2026
GitHub β†—β˜… 9.1Kβ‘‚ 778
OCR & Documents

DocumindHQ/documind

Open-source platform for extracting structured data from documents using AI.

JavaScriptOtherupdated May 15, 2025
GitHub β†—β˜… 1.5Kβ‘‚ 62
OCR & Documents

pymupdf/PyMuPDFπŸ”₯ active

PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.

PythonAGPL-3.0updated Sep 24, 2026
GitHub β†—β˜… 10.8Kβ‘‚ 804
OCR & Documents

PaddlePaddle/PaddleOCR

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

PythonApache-2.0updated Sep 16, 2026
GitHub β†—β˜… 90.1Kβ‘‚ 11.4K
OCR & Documents

ocrmypdf/OCRmyPDFπŸ”₯ active

OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched

PythonMPL-2.0updated Sep 22, 2026
GitHub β†—β˜… 34.9Kβ‘‚ 2.4K

Data from GitHub Β· snapshot Sep 24, 2026