PKU-Alignment/safe-rlhf
Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
Load local LLMs effortlessly in a Jupyter notebook for testing purposes alongside Langchain or other agents. Contains Oobagooga and KoboldAI versions of the langchain notebooks with examples.
$ git clone https://github.com/ausboss/Local-LLM-Langchain.gitSafe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
A complete guide to start and improve your LLM skills in 2026 with little background in the field and stay up-to-date with the latest news and state-of-the-art techniques!
中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)
中文LLaMA-2 & Alpaca-2大模型二期项目 + 64K超长上下文模型 (Chinese LLaMA-2 & Alpaca-2 LLMs with 64K long context models)
General technology for enabling AI capabilities w/ LLMs and MLLMs
Build a modern LLM from scratch. Every line commented. Explained like we are five.
Data from GitHub · snapshot Sep 24, 2026