dair-ai/Prompt-Engineering-Guide
π Guides, papers, lessons, notebooks and resources for prompt engineering, context engineering, RAG, and AI Agents.
[ICLR 2025] A trinity of environments, tools, and benchmarks for general virtual agents
$ git clone https://github.com/ltzheng/agent-studio.gitπ Guides, papers, lessons, notebooks and resources for prompt engineering, context engineering, RAG, and AI Agents.
π‘ All-in-one AI framework for semantic search, LLM orchestration and language model workflows
SWE-bench: Can Language Models Resolve Real-world Github Issues?
δΈζθ―θ¨ηθ§£ζ΅θ―εΊε Chinese Language Understanding Evaluation Benchmark: datasets, baselines, pre-trained models, corpus and leaderboard
[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments
The only fully local production-grade Super SDK that provides a simple, unified, and powerful interface for calling more than 200+ LLMs.
Data from GitHub Β· snapshot Sep 24, 2026