๐Ÿ† #1,769 overall#270 of 313 in OCR & Documents

AlexanderPro /WindowTextExtractor

WindowTextExtractor allows you to get a text from any window of an operating system including asterisk passwords

$ git clone https://github.com/AlexanderPro/WindowTextExtractor.git
GitHub social preview for AlexanderPro/WindowTextExtractor
Stars
396
396
Forks
36
36
Language
C#
License
MIT
Created
Oct 27, 2019
6.9 years old
Last push
Oct 4, 2025

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

YaoFANGUK/video-subtitle-extractor

่ง†้ข‘็กฌๅญ—ๅน•ๆๅ–๏ผŒ็”Ÿๆˆsrtๆ–‡ไปถใ€‚ๆ— ้œ€็”ณ่ฏท็ฌฌไธ‰ๆ–นAPI๏ผŒๆœฌๅœฐๅฎž็Žฐๆ–‡ๆœฌ่ฏ†ๅˆซใ€‚ๅŸบไบŽๆทฑๅบฆๅญฆไน ็š„่ง†้ข‘ๅญ—ๅน•ๆๅ–ๆก†ๆžถ๏ผŒๅŒ…ๅซๅญ—ๅน•ๅŒบๅŸŸๆฃ€ๆต‹ใ€ๅญ—ๅน•ๅ†…ๅฎนๆๅ–ใ€‚A GUI tool for extracting hard-coded subtitle (hardsub) from videos and generating srt files.

PythonApache-2.0updated Apr 9, 2026
GitHub โ†—โ˜… 9.5Kโ‘‚ 969
OCR & Documents

CatchTheTornado/text-extract-api

Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported models. Anonymize documents. Remove PII. Convert any document or picture to structured JSON or Markdown

PythonMITupdated Dec 8, 2025
GitHub โ†—โ˜… 3.2Kโ‘‚ 279
OCR & Documents

datalab-to/lift

Extract structured data from documents quickly and accurately.

PythonApache-2.0updated Jun 19, 2026
GitHub โ†—โ˜… 910โ‘‚ 85
OCR & Documents

PaddlePaddle/PaddleOCR

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

PythonApache-2.0updated Sep 16, 2026
GitHub โ†—โ˜… 90.1Kโ‘‚ 11.4K
OCR & Documents

opendatalab/MinerU๐Ÿ”ฅ active

Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.

PythonOtherupdated Sep 24, 2026
GitHub โ†—โ˜… 80.6Kโ‘‚ 6.7K

Data from GitHub ยท snapshot Sep 24, 2026