OptimalScale/LMFlow
An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.
An implementation of DeepMind's Relational Recurrent Neural Networks (NeurIPS 2018) in PyTorch.
$ git clone https://github.com/L0SG/relational-rnn-pytorch.gitAn Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.
ExtremeBERT is a toolkit that accelerates the pretraining of customized language models on customized datasets, described in the paper βExtremeBERT: A Toolkit for Accelerating Pretraining of Customized BERTβ.
Source code for "Learning protein sequence embeddings using information from structure" - ICLR 2019
RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.
Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
Data from GitHub Β· snapshot Sep 24, 2026