filyp/autocorrect
Spelling corrector in python
An local, offline (after initial setup), portable OCR software that can process images and PDF files, using DeepSeek-OCR-2 AI (running directly on your machine).
$ git clone https://github.com/th1nhhdk/local_ai_ocr.gitSpelling corrector in python
A community-supported supercharged document management system: scan, index and archive all your documents
PDF craft can convert PDF files into various other formats. This project will focus on processing PDF files of scanned books.
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
Fast and efficient unstructured data extraction. Written in Rust with bindings for many languages.
ExtractThinker is a Document Intelligence library for LLMs, offering ORM-style interaction for flexible and powerful document workflows.
Data from GitHub ยท snapshot Sep 24, 2026