OCR & Documents
naptha/tesseract.js
Pure Javascript OCR for more than 100 Languages ๐๐๐ฅ
JavaScriptApache-2.0updated May 17, 2026
JavaScript optical character recognition demo
$ git clone https://github.com/kdzwinel/JS-OCR-demo.gitPure Javascript OCR for more than 100 Languages ๐๐๐ฅ
Lightweight document management system packed with all the features you can expect from big expensive solutions
Chrome extension: pin a screen region once, then hotkey your way through a paginated document. OCR runs 100% offline via bundled Tesseract.
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
Data from GitHub ยท snapshot Sep 24, 2026