raiyanyahya/how-to-train-your-gpt
Build a modern LLM from scratch. Every line commented. Explained like we are five.
Bayesian Optimization as a Coverage Tool for Evaluating LLMs. Accurate evaluation (benchmarking) that's 10 times faster with just a few lines of modular code.
$ git clone https://github.com/rentruewang/bocoel.gitBuild a modern LLM from scratch. Every line commented. Explained like we are five.
Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard and designing lighteval!
The official implementation of RAPTOR: Recursive Abstractive Processing for Tree-Organized Retrieval
INT4/INT5/INT8 and FP16 inference on CPU for RWKV language model
Grounded search engine (i.e. with source reference) based on LLM / ChatGPT / OpenAI API. It supports web search, file content search etc.
This repo contains evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI"
Data from GitHub ยท snapshot Sep 24, 2026