๐Ÿ† #1,117 overall#164 of 313 in OCR & Documents

yigitkonur /api-llm-ocr

PDF to markdown using vision LLMs โ€” tables, layouts, and structure preserved

$ git clone https://github.com/yigitkonur/api-llm-ocr.git
GitHub social preview for yigitkonur/api-llm-ocr
Stars
901
901
Forks
61
61
Language
Python
License
Other
Created
Sep 22, 2024
2.0 years old
Last push
Feb 21, 2026

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

aiptimizer/TurboOCR

TurboOCR, >200 img/s OmnidocBench. TensorRT FP16, PP-OCRv6, HTTP + gRPC

C++MITupdated Sep 8, 2026
GitHub โ†—โ˜… 1.1Kโ‘‚ 112
OCR & Documents

run-llama/ParseBench๐Ÿ”ฅ active

ParseBench - A Document Parsing Benchmark for AI Agents

PythonApache-2.0updated Sep 23, 2026
GitHub โ†—โ˜… 592โ‘‚ 109
OCR & Documents

neosun100/DeepSeek-OCR-WebUI

๐ŸŽจ Ready-to-use DeepSeek-OCR Web UI | Modern Interface | 7 Recognition Modes | Batch Processing | Real-time Logging | Fully Responsive

TypeScriptMITupdated Feb 20, 2026
GitHub โ†—โ˜… 447โ‘‚ 95
OCR & Documents

ocrmypdf/OCRmyPDF๐Ÿ”ฅ active

OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched

PythonMPL-2.0updated Sep 22, 2026
GitHub โ†—โ˜… 34.9Kโ‘‚ 2.4K
OCR & Documents

run-llama/liteparse๐Ÿ”ฅ active

A fast, helpful, and open-source document parser

RustApache-2.0updated Sep 22, 2026
GitHub โ†—โ˜… 12.6Kโ‘‚ 860
OCR & Documents

pymupdf/PyMuPDF๐Ÿ”ฅ active

PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.

PythonAGPL-3.0updated Sep 24, 2026
GitHub โ†—โ˜… 10.8Kโ‘‚ 804

Data from GitHub ยท snapshot Sep 24, 2026