OCR & Documents
zhoubear/open-paperless
Scan, index, and archive all of your paper documents (acquired by Mayan EDMS)
PythonOtherupdated Dec 10, 2018
I, Librarian - open-source version of a PDF managing SaaS.
$ git clone https://github.com/mkucej/i-librarian-free.gitScan, index, and archive all of your paper documents (acquired by Mayan EDMS)
Assist in organizing your piles of documents, resulting from scanners, e-mails and other sources with miminal effort.
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
A community-supported supercharged document management system: scan, index and archive all your documents
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
A fast, helpful, and open-source document parser
Data from GitHub ยท snapshot Sep 24, 2026