๐Ÿ† #1,834 overall#282 of 313 in OCR & Documents

harishdeivanayagam /rowfill

Open-source spreadsheets platform for deep research and document processing

$ git clone https://github.com/harishdeivanayagam/rowfill.git
GitHub social preview for harishdeivanayagam/rowfill
Stars
367
367
Forks
21
21
Language
TypeScript
License
Other
Created
Jan 9, 2025
1.7 years old
Last push
Sep 25, 2025

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

landing-ai/ade-cli

The official CLI for Agentic Document Extraction (ADE) by LandingAI โ€” parse documents and extract schema-shaped data from your terminal

PythonApache-2.0updated Aug 19, 2026
GitHub โ†—โ˜… 2.4Kโ‘‚ 255
OCR & Documents

enoch3712/ExtractThinker

ExtractThinker is a Document Intelligence library for LLMs, offering ORM-style interaction for flexible and powerful document workflows.

๐Ÿง  Natural Language Processing
PythonApache-2.0updated Sep 16, 2026
GitHub โ†—โ˜… 1.6Kโ‘‚ 153
OCR & Documents

NanoNets/docstrange

Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats (Markdown, JSON, CSV, HTML) with intelligent structured data extraction and advanced OCR.

PythonMITupdated Oct 31, 2025
GitHub โ†—โ˜… 1.6Kโ‘‚ 139
OCR & Documents

run-llama/ParseBench๐Ÿ”ฅ active

ParseBench - A Document Parsing Benchmark for AI Agents

PythonApache-2.0updated Sep 23, 2026
GitHub โ†—โ˜… 592โ‘‚ 109
OCR & Documents

PaddlePaddle/PaddleOCR

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

PythonApache-2.0updated Sep 16, 2026
GitHub โ†—โ˜… 90.1Kโ‘‚ 11.4K
OCR & Documents

paperless-ngx/paperless-ngx๐Ÿ”ฅ active

A community-supported supercharged document management system: scan, index and archive all your documents

PythonGPL-3.0updated Sep 24, 2026
GitHub โ†—โ˜… 46Kโ‘‚ 3.2K

Data from GitHub ยท snapshot Sep 24, 2026