πŸ† #1,152 overall#333 of 601 in Language Models

ethz-spylab /agentdojo

A Dynamic Environment to Evaluate Attacks and Defenses for LLM Agents.

$ git clone https://github.com/ethz-spylab/agentdojo.git
GitHub social preview for ethz-spylab/agentdojo
Stars
865
865
Forks
221
221
Language
Python
License
MIT
Created
Feb 29, 2024
2.6 years old
Last push
Jun 2, 2026

Categories

GitHub topics

More in Language Models

Language Models

EvolvingLMMs-Lab/lmms-evalπŸ”₯ active

One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks

PythonOtherupdated Sep 24, 2026
GitHub β†—β˜… 4.4Kβ‘‚ 658
Language Models

baichuan-inc/Baichuan2

A series of large language models developed by Baichuan Intelligent Technology

🧠 Natural Language Processing
PythonApache-2.0updated Nov 8, 2024
GitHub β†—β˜… 4.1Kβ‘‚ 289
Language Models

RUC-NLPIR/FlashRAGπŸ”₯ active

⚑FlashRAG: A Python Toolkit for Efficient RAG Research (WWW2025 Resource)

PythonMITupdated Sep 19, 2026
GitHub β†—β˜… 3.6Kβ‘‚ 315
Language Models

baichuan-inc/Baichuan-13B

A 13B large language model developed by Baichuan Intelligent Technology

🧠 Natural Language Processing
PythonApache-2.0updated Sep 6, 2023
GitHub β†—β˜… 2.9Kβ‘‚ 227
Language Models

evalplus/evalplus

Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024

PythonApache-2.0updated Oct 2, 2025
GitHub β†—β˜… 1.8Kβ‘‚ 208
Language Models

MLGroupJLU/LLM-eval-survey

The official GitHub page for the survey paper "A Survey on Evaluation of Large Language Models".

Otherno licenseupdated Sep 13, 2026
GitHub β†—β˜… 1.6Kβ‘‚ 105

Data from GitHub Β· snapshot Sep 24, 2026