๐Ÿ† #555 overall#170 of 601 in Language Models

dvmazur /mixtral-offloading

Run Mixtral-8x7B models in Colab or consumer desktops

$ git clone https://github.com/dvmazur/mixtral-offloading.git
GitHub social preview for dvmazur/mixtral-offloading
Stars
2.3K
2,335
Forks
225
225
Language
Python
License
MIT
Created
Dec 15, 2023
2.8 years old
Last push
Apr 8, 2024

Categories

GitHub topics

More in Language Models

Language Models

RWKV/rwkv.cpp

INT4/INT5/INT8 and FP16 inference on CPU for RWKV language model

C++MITupdated Mar 23, 2025
GitHub โ†—โ˜… 1.6Kโ‘‚ 130
Language Models

RahulSChand/gpu_poor

Calculate token/s & GPU memory requirement for any LLM. Supports llama.cpp/ggml/bnb/QLoRA quantization

JavaScriptno licenseupdated Dec 3, 2024
GitHub โ†—โ˜… 1.4Kโ‘‚ 88
Language Models

AviSoori1x/makeMoE

From scratch implementation of a sparse mixture of experts language model inspired by Andrej Karpathy's makemore :)

Jupyter NotebookMITupdated Oct 30, 2024
GitHub โ†—โ˜… 816โ‘‚ 96
Language Models

mlc-ai/web-llm

High-performance In-browser LLM Inference Engine

TypeScriptApache-2.0updated Sep 15, 2026
GitHub โ†—โ˜… 19.2Kโ‘‚ 1.4K
Language Models

BlinkDL/RWKV-LM๐Ÿ”ฅ active

RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.

PythonApache-2.0updated Sep 21, 2026
GitHub โ†—โ˜… 14.7Kโ‘‚ 1K

Data from GitHub ยท snapshot Sep 24, 2026