Language Models
SWE-bench/SWE-smith๐ฅ active
[NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agents
PythonMITupdated Sep 21, 2026
Benchmarking Goal-Oriented Software Engineering
$ git clone https://github.com/CodeClash-ai/CodeClash.git[NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agents
๐ Guides, papers, lessons, notebooks and resources for prompt engineering, context engineering, RAG, and AI Agents.
Private AI platform for agents, assistants and enterprise search. Built-in Agent Builder, Deep research, Document analysis, Multi-model support, and API connectivity for agents.
๐ก All-in-one AI framework for semantic search, LLM orchestration and language model workflows
SWE-bench: Can Language Models Resolve Real-world Github Issues?
Harness LLMs with Multi-Agent Programming
Data from GitHub ยท snapshot Sep 24, 2026