🏆 #2,344 overall#88 of 158 in Information Extraction

ck-unifr /pdf_parsing

PDF解析(文字,章节,表格,图片,参考),基于大模型(ChatGLM2-6B, RWKV)+langchain+streamlit的PDF问答,摘要,信息抽取

$ git clone https://github.com/ck-unifr/pdf_parsing.git
GitHub social preview for ck-unifr/pdf_parsing
Stars
215
215
Forks
30
30
Language
Python
License
None
Created
Sep 8, 2023
3.0 years old
Last push
Oct 17, 2023

Categories

GitHub topics

More in Information Extraction

Information Extraction

yifanfeng97/Hyper-Extract🔥 active

Hypergraph is more powerful. Transform unstructured text into structured knowledge with LLMs. Graphs, hypergraphs, and spatio-temporal extractions — with one command.

PythonOtherupdated Sep 20, 2026
GitHub ↗★ 4K⑂ 457
Information Extraction

langstruct-ai/langstruct

Extract structured data from any content using LLMs.

🤖 Language Models🧠 Natural Language Processing
PythonMITupdated Dec 1, 2025
GitHub ↗★ 130⑂ 13
Information Extraction

PaddlePaddle/PaddleNLP

Easy-to-use and powerful LLM and SLM library with awesome model zoo.

💖 Sentiment Analysis❓ Question Answering🧠 Natural Language Processing
PythonApache-2.0updated May 23, 2026
GitHub ↗★ 13K⑂ 3K
Information Extraction

mit-nlp/MITIE

MITIE: library and tools for information extraction

🧠 Natural Language Processing
C++no licenseupdated Sep 28, 2025
GitHub ↗★ 3K⑂ 532
Information Extraction

eyurtsev/kor

LLM(😽)

🧠 Natural Language Processing
PythonMITupdated Feb 3, 2025
GitHub ↗★ 1.7K⑂ 93
Information Extraction

monarch-initiative/ontogpt

LLM-based ontological extraction tools, including SPIRES

🏷️ Named Entity Recognition🤖 Language Models🧠 Natural Language Processing
Jupyter NotebookBSD-3-Clauseupdated Sep 10, 2026
GitHub ↗★ 1K⑂ 124

Data from GitHub · snapshot Sep 24, 2026