jshuadvd/LongRoPE
Implementation of the LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens Paper
Trankit is a Light-Weight Transformer-based Python Toolkit for Multilingual Natural Language Processing
$ git clone https://github.com/nlp-uoregon/trankit.gitImplementation of the LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens Paper
๐ซ Industrial-strength Natural Language Processing (NLP) in Python
๐ spaCy building blocks and visualizers for Streamlit apps
CogComp's Natural Language Processing Libraries and Demos: Modules include lemmatizer, ner, pos, prep-srl, quantifier, question type, relation-extraction, similarity, temporal normalizer, tokenizer, transliteration, verb-sense, and more.
NLP Cheat Sheet, Python, spacy, LexNPL, NLTK, tokenization, stemming, sentence detection, named entity recognition
[LREC 2022] An off-the-shelf pre-trained Tweet NLP Toolkit (NER, tokenization, lemmatization, POS tagging, dependency parsing) + Tweebank-NER dataset
Data from GitHub ยท snapshot Sep 24, 2026