bytedance/Dolphin
The official repo for βDolphin: Document Image Parsing via Heterogeneous Anchor Promptingβ, ACL, 2025.
A fast, helpful, and open-source document parser
$ git clone https://github.com/run-llama/liteparse.gitThe official repo for βDolphin: Document Image Parsing via Heterogeneous Anchor Promptingβ, ACL, 2025.
ε¨δΏηηι’γε ¬εΌδΈη»ζηεζδΈθΏθ‘ PDF ηΏ»θ―οΌιη¨δΊη§η δΈζζ―ζζ‘£
Citra β PDF answers with page-level proof. Local-first structured text, tables, OCR, visual evidence, and citations via MCP, CLI, and SDK.
Java PDF table extraction & OCR library. Extract structured tables from text-based and scanned PDFs using stream, lattice (OpenCV-style grid detection), and hybrid parsing.
Open-source batch OCR workbench β a free, local alternative to ABBYY FineReader. Powered by Ollama + GLM-OCR + PP-DocLayoutV3, ~0.5s/page on RTX 4090. Three-panel editor, layout-aware, PDF/image batch processing, Markdown/Word export. ζΉιOCRε·₯δ½ε°οΌηΊ―ζ¬ε°θΏθ‘οΌε θ΄ΉεΉ³ζΏABBYYοΌιεδΉ¦η±ζζ‘£ζ°εεγ
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
Data from GitHub Β· snapshot Sep 24, 2026