πŸ† #1,472 overall#220 of 313 in OCR & Documents

wenwenyu /PICK-pytorch

Code for the paper "PICK: Processing Key Information Extraction from Documents using Improved Graph Learning-Convolutional Networks" (ICPR 2020)

$ git clone https://github.com/wenwenyu/PICK-pytorch.git
GitHub social preview for wenwenyu/PICK-pytorch
Stars
570
570
Forks
189
189
Language
Python
License
MIT
Created
Jul 15, 2020
6.2 years old
Last push
Jul 25, 2024

Categories

GitHub topics

More in OCR & Documents

OCR & Documents

AlibabaResearch/AdvancedLiterateMachinery

A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team in the Language Technology Lab, Tongyi Lab, Alibaba Group.

C++Apache-2.0updated Mar 17, 2026
GitHub β†—β˜… 1.8Kβ‘‚ 195
OCR & Documents

jpWang/LiLT

Official PyTorch implementation of LiLT: A Simple yet Effective Language-Independent Layout Transformer for Structured Document Understanding (ACL 2022)

⛏️ Information Extraction🧠 Natural Language Processing
PythonMITupdated Oct 31, 2022
GitHub β†—β˜… 370β‘‚ 40
OCR & Documents

andreagemelli/doc2graph

Doc2Graph transforms documents into graphs and exploit a GNN to solve several tasks.

🧠 Natural Language Processing
Jupyter NotebookMITupdated Oct 18, 2025
GitHub β†—β˜… 142β‘‚ 25
OCR & Documents

opendatalab/MinerUπŸ”₯ active

Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.

PythonOtherupdated Sep 24, 2026
GitHub β†—β˜… 80.6Kβ‘‚ 6.7K
OCR & Documents

bytedance/Dolphin

The official repo for β€œDolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.

PythonOtherupdated Mar 25, 2026
GitHub β†—β˜… 9.1Kβ‘‚ 778

Data from GitHub Β· snapshot Sep 24, 2026