πŸ† #1,778 overall#447 of 601 in Language Models

FranxYao /Deep-Generative-Models-for-Natural-Language-Processing

DGMs for NLP. A roadmap.

$ git clone https://github.com/FranxYao/Deep-Generative-Models-for-Natural-Language-Processing.git
GitHub social preview for FranxYao/Deep-Generative-Models-for-Natural-Language-Processing
Stars
393
393
Forks
33
33
Language
Other
License
None
Created
Apr 9, 2019
7.5 years old
Last push
Dec 12, 2022

Categories

GitHub topics

More in Language Models

Language Models

salesforce/progen

Official release of the ProGen models

PythonBSD-3-Clauseupdated Jun 2, 2026
GitHub β†—β˜… 705β‘‚ 139
Language Models

rasbt/LLMs-from-scratchπŸ”₯ active

Implement a ChatGPT-like LLM in PyTorch from scratch, step by step

βœ‚οΈ Tokenization & Preprocessing🧠 Natural Language Processing
Jupyter NotebookOtherupdated Sep 22, 2026
GitHub β†—β˜… 105.5Kβ‘‚ 16.2K
Language Models

arc53/DocsGPTπŸ”₯ active

Private AI platform for agents, assistants and enterprise search. Built-in Agent Builder, Deep research, Document analysis, Multi-model support, and API connectivity for agents.

πŸ” Search & Retrieval🧠 Natural Language Processing
PythonMITupdated Sep 23, 2026
GitHub β†—β˜… 18.3Kβ‘‚ 2.2K
Language Models

neuml/txtaiπŸ”₯ active

πŸ’‘ All-in-one AI framework for semantic search, LLM orchestration and language model workflows

🧬 EmbeddingsπŸ” Search & Retrieval🧠 Natural Language Processing
PythonApache-2.0updated Sep 23, 2026
GitHub β†—β˜… 13Kβ‘‚ 899
Language Models

huggingface/tokenizersπŸ”₯ active

πŸ’₯ Fast State-of-the-Art Tokenizers optimized for Research and Production

🧠 Natural Language Processing
RustApache-2.0updated Sep 23, 2026
GitHub β†—β˜… 11.1Kβ‘‚ 1.2K
Language Models

brightmart/nlp_chinese_corpus

倧规樑中文θ‡ͺ焢语言倄理语料 Large Scale Chinese Corpus for NLP

πŸ—‚οΈ Text Classification❓ Question Answering🧬 Embeddings🧠 Natural Language Processing
OtherMITupdated Feb 6, 2026
GitHub β†—β˜… 9.9Kβ‘‚ 1.6K

Data from GitHub Β· snapshot Sep 24, 2026