PKU-Alignment/safe-rlhf
Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
BeaverTails is a collection of datasets designed to facilitate research on safety alignment in large language models (LLMs).
$ git clone https://github.com/PKU-Alignment/beavertails.gitSafe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
Build a modern LLM from scratch. Every line commented. Explained like we are five.
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
A complete guide to start and improve your LLM skills in 2026 with little background in the field and stay up-to-date with the latest news and state-of-the-art techniques!
[COLM 2024] OpenAgents: An Open Platform for Language Agents in the Wild
General technology for enabling AI capabilities w/ LLMs and MLLMs
Data from GitHub ยท snapshot Sep 24, 2026