🏆 #752 overall#216 of 601 in Language Models

PKU-Alignment /safe-rlhf

Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback

$ git clone https://github.com/PKU-Alignment/safe-rlhf.git
GitHub social preview for PKU-Alignment/safe-rlhf
Stars
1.6K
1,618
Forks
134
134
Language
Python
License
Apache-2.0
Created
May 15, 2023
3.4 years old
Last push
Nov 24, 2025

Categories

GitHub topics

More in Language Models

Language Models

PKU-Alignment/beavertails

BeaverTails is a collection of datasets designed to facilitate research on safety alignment in large language models (LLMs).

MakefileApache-2.0updated Oct 27, 2023
GitHub ↗★ 183⑂ 6
Language Models

louisfb01/start-llms

A complete guide to start and improve your LLM skills in 2026 with little background in the field and stay up-to-date with the latest news and state-of-the-art techniques!

OtherMITupdated Jan 23, 2026
GitHub ↗★ 983⑂ 128
Language Models

ymcui/Chinese-LLaMA-Alpaca

中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)

🧠 Natural Language Processing
PythonApache-2.0updated Apr 19, 2026
GitHub ↗★ 18.9K⑂ 1.8K
Language Models

ymcui/Chinese-LLaMA-Alpaca-2

中文LLaMA-2 & Alpaca-2大模型二期项目 + 64K超长上下文模型 (Chinese LLaMA-2 & Alpaca-2 LLMs with 64K long context models)

🧠 Natural Language Processing
PythonApache-2.0updated Apr 19, 2026
GitHub ↗★ 7.1K⑂ 558
Language Models

darrenburns/elia

A snappy, keyboard-centric terminal user interface for interacting with large language models. Chat with ChatGPT, Claude, Llama 3, Phi 3, Mistral, Gemma and more.

PythonApache-2.0updated Oct 10, 2024
GitHub ↗★ 2.5K⑂ 154
Language Models

ymcui/Chinese-LLaMA-Alpaca-3

中文羊驼大模型三期项目 (Chinese Llama-3 LLMs) developed from Meta Llama 3

🧠 Natural Language Processing
PythonApache-2.0updated Apr 19, 2026
GitHub ↗★ 2K⑂ 169

Data from GitHub · snapshot Sep 24, 2026