🏆 #1,668 overall#433 of 601 in Language Models

Joyce94 /LLM-RLHF-Tuning

LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)

$ git clone https://github.com/Joyce94/LLM-RLHF-Tuning.git
GitHub social preview for Joyce94/LLM-RLHF-Tuning
Stars
452
452
Forks
24
24
Language
Python
License
None
Created
Jun 12, 2023
3.3 years old
Last push
Oct 11, 2023

Categories

GitHub topics

More in Language Models

Language Models

louisfb01/start-llms

A complete guide to start and improve your LLM skills in 2026 with little background in the field and stay up-to-date with the latest news and state-of-the-art techniques!

OtherMITupdated Jan 23, 2026
GitHub ↗★ 983⑂ 128
Language Models

WangRongsheng/Aurora

The official codes for "Aurora: Activating chinese chat capability for Mixtral-8x7B sparse Mixture-of-Experts through Instruction-Tuning"

PythonApache-2.0updated May 9, 2024
GitHub ↗★ 260⑂ 19
Language Models

ymcui/Chinese-LLaMA-Alpaca

中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)

🧠 Natural Language Processing
PythonApache-2.0updated Apr 19, 2026
GitHub ↗★ 18.9K⑂ 1.8K
Language Models

raiyanyahya/how-to-train-your-gpt

Build a modern LLM from scratch. Every line commented. Explained like we are five.

🧠 Natural Language Processing
Jupyter NotebookMITupdated Aug 30, 2026
GitHub ↗★ 3.4K⑂ 419
Language Models

stochasticai/xTuring

Build, personalize and control your own LLMs. From data pre-processing to fine-tuning, xTuring provides an easy way to personalize open-source LLMs. Join our discord community: https://discord.gg/TgHXuSJEk6

PythonApache-2.0updated Sep 12, 2026
GitHub ↗★ 2.7K⑂ 210
Language Models

RahulSChand/gpu_poor

Calculate token/s & GPU memory requirement for any LLM. Supports llama.cpp/ggml/bnb/QLoRA quantization

JavaScriptno licenseupdated Dec 3, 2024
GitHub ↗★ 1.4K⑂ 88

Data from GitHub · snapshot Sep 24, 2026