OCR & Documents
scambier/obsidian-text-extractor
A (companion) plugin to facilitate the extraction of text from images (OCR) and PDFs.
TypeScriptGPL-3.0updated Sep 7, 2026
A search engine that "just works" for Obsidian. Supports OCR and PDF indexing.
$ git clone https://github.com/scambier/obsidian-omnisearch.gitA (companion) plugin to facilitate the extraction of text from images (OCR) and PDFs.
Extract annotations (highlights and scribbles) from PDF, EPUB, and notebooks marked with reMarkable tablets. Export to Markdown, PDF, PNG, SVG
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
A fast, helpful, and open-source document parser
Data from GitHub ยท snapshot Sep 24, 2026