๐Ÿ† #1,101 overall#161 of 313 in OCR & Documents

eclaire-labs /eclaire

Local-first, open-source AI assistant for your data. Unify tasks, notes, docs, photos, and bookmarks. Private, self-hosted, and extensible via APIs.

$ git clone https://github.com/eclaire-labs/eclaire.git
GitHub social preview for eclaire-labs/eclaire
Stars
922
922
Forks
96
96
Language
TypeScript
License
MIT
Created
Sep 2, 2025
1.1 years old
Last push
May 14, 2026

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

enoch3712/ExtractThinker

ExtractThinker is a Document Intelligence library for LLMs, offering ORM-style interaction for flexible and powerful document workflows.

๐Ÿง  Natural Language Processing
PythonApache-2.0updated Sep 16, 2026
GitHub โ†—โ˜… 1.6Kโ‘‚ 153
OCR & Documents

jztan/pdf-mcp๐Ÿ”ฅ active

MCP server that lets Claude Code and other AI agents read and search large PDFs, one file or a whole folder: agentic RAG with hybrid semantic + keyword search, selective page reads, tables, images, OCR, chart data, and multi-column/CJK layouts.

๐Ÿ” Search & Retrieval
PythonMITupdated Sep 22, 2026
GitHub โ†—โ˜… 139โ‘‚ 14
OCR & Documents

paperless-ngx/paperless-ngx๐Ÿ”ฅ active

A community-supported supercharged document management system: scan, index and archive all your documents

PythonGPL-3.0updated Sep 24, 2026
GitHub โ†—โ˜… 46Kโ‘‚ 3.2K
OCR & Documents

dataelement/bisheng๐Ÿ”ฅ active

BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehensive features include: GenAI workflow, RAG, Agent, Unified model management, Evaluation, SFT, Dataset Management, Enterprise-level System Management, Observability and more.

PythonApache-2.0updated Sep 24, 2026
GitHub โ†—โ˜… 12Kโ‘‚ 2K
OCR & Documents

icereed/paperless-gpt๐Ÿ”ฅ active

Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI

GoMITupdated Sep 24, 2026
GitHub โ†—โ˜… 2.7Kโ‘‚ 211
OCR & Documents

NanoNets/docstrange

Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats (Markdown, JSON, CSV, HTML) with intelligent structured data extraction and advanced OCR.

PythonMITupdated Oct 31, 2025
GitHub โ†—โ˜… 1.6Kโ‘‚ 139

Data from GitHub ยท snapshot Sep 24, 2026