Language Models
langfuse/langfuse๐ฅ active
๐ชข Open source agent evals & observability: Trace, evaluate, and improve LLM applications with one open platform.
TypeScriptOtherupdated Sep 24, 2026
An open-source visual programming environment for battle-testing prompts to LLMs.
$ git clone https://github.com/ianarawjo/ChainForge.git๐ชข Open source agent evals & observability: Trace, evaluate, and improve LLM applications with one open platform.
20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
Implement a reasoning LLM in PyTorch from scratch, step by step
Open-source tools for prompt testing and experimentation, with support for both LLMs (e.g. OpenAI, LLaMA) and vector databases (e.g. Chroma, Weaviate, LanceDB).
The official GitHub page for the survey paper "A Survey on Evaluation of Large Language Models".
Data from GitHub ยท snapshot Sep 24, 2026