OCR & Documents
Layout-Parser/layout-parser
A Unified Toolkit for Deep Learning Based Document Image Analysis
PythonApache-2.0updated Aug 15, 2024
All-in-One Development Tool based on PaddlePaddle
$ git clone https://github.com/PaddlePaddle/PaddleX.gitA Unified Toolkit for Deep Learning Based Document Image Analysis
Machine Learning Training Utilities (for TensorFlow and PyTorch)
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
High-performance, privacy-first on-device AI inference library for React Native, powered by PyTorch's ExecuTorch runtime
CnSTD: 基于 PyTorch/MXNet 的 中文/英文 场景文字检测(Scene Text Detection)、数学公式检测(Mathematical Formula Detection, MFD)、篇章分析(Layout Analysis)的Python3 包
Data from GitHub · snapshot Sep 24, 2026