πŸ† #1,760 overall#267 of 313 in OCR & Documents

lucasrla /remarks

Extract annotations (highlights and scribbles) from PDF, EPUB, and notebooks marked with reMarkable tablets. Export to Markdown, PDF, PNG, SVG

$ git clone https://github.com/lucasrla/remarks.git
GitHub social preview for lucasrla/remarks
Stars
399
399
Forks
34
34
Language
Python
License
GPL-3.0
Created
Jul 26, 2020
6.2 years old
Last push
May 26, 2024

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

pymupdf/PyMuPDFπŸ”₯ active

PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.

PythonAGPL-3.0updated Sep 24, 2026
GitHub β†—β˜… 10.8Kβ‘‚ 804
OCR & Documents

opendatalab/MinerUπŸ”₯ active

Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.

PythonOtherupdated Sep 24, 2026
GitHub β†—β˜… 80.6Kβ‘‚ 6.7K
OCR & Documents

bytedance/Dolphin

The official repo for β€œDolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.

PythonOtherupdated Mar 25, 2026
GitHub β†—β˜… 9.1Kβ‘‚ 778
OCR & Documents

oomol-lab/pdf-craftπŸ”₯ active

PDF craft can convert PDF files into various other formats. This project will focus on processing PDF files of scanned books.

PythonMITupdated Sep 23, 2026
GitHub β†—β˜… 6.3Kβ‘‚ 458
OCR & Documents

scambier/obsidian-omnisearchπŸ”₯ active

A search engine that "just works" for Obsidian. Supports OCR and PDF indexing.

TypeScriptGPL-3.0updated Sep 19, 2026
GitHub β†—β˜… 2.2Kβ‘‚ 118
OCR & Documents

dengxibo/sumatrapdf-plusπŸ”₯ active

SumatraPDF fork: Chinese EPUB/MOBI, smart PDF dark mode, OCR, TTS, offline dictionary.

CGPL-3.0updated Sep 24, 2026
GitHub β†—β˜… 862β‘‚ 32

Data from GitHub Β· snapshot Sep 24, 2026