yigitkonur/api-llm-ocr
PDF to markdown using vision LLMs โ tables, layouts, and structure preserved
TurboOCR, >200 img/s OmnidocBench. TensorRT FP16, PP-OCRv6, HTTP + gRPC
$ git clone https://github.com/aiptimizer/TurboOCR.gitPDF to markdown using vision LLMs โ tables, layouts, and structure preserved
ParseBench - A Document Parsing Benchmark for AI Agents
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.
PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
๐ Awesome OCR multiple programing languages toolkits based on ONNX Runtime, OpenVINO, MNN, PaddlePaddle, TensorRT and PyTorch.
Data from GitHub ยท snapshot Sep 24, 2026