// OWNER · THUDM · REPO SNAPSHOT
THUDM
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
STARS3,603+18 · 7d
FORKS272+0 · 7d
CONTRIBUTORS0maintainer graph thin
OPEN ISSUES · PRS63· 0
LICENSE—no license file
LANGUAGEPython
FOUNDED—
LOCATION—
MEMBERS—
PUBLIC REPOS—
TOTAL STARS—
FOCUS—
FUNDING—
//STAR HISTORY · 12M
+115 all-time-window
Stars · cumulative3,603 · today
//WHY IT'S TRENDING· narrative synthesis · medium confidence
THUDM/AgentBench is sitting at #2854 on the trending leaderboard with a pulse of 2/100 with no cross-source channels firing yet — GitHub-stars-only signal so far.
The 7-day star delta is +18 (+0.5%) against a base of 3,603, so the headline number is category momentum more than a personal breakout. What's actually moving is the AI agent / LLM tooling stack.
Watch-outs: no tagged release on record (treat as pre-stable).
↻ refreshed 30m ago· next sync in 35m· sourced from 0 channels · 0 mentions / 24h
// RELATED REPOS · CROSS-SOURCE OVERLAP
▌ FAQ · answers from the data spine
What is THUDM/AgentBench?
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
Who maintains AgentBench?
Maintained by THUDM (the owner) on GitHub.
What language is AgentBench written in?
Primarily Python.
When was AgentBench created? When was the last update?
Created July 28, 2023 · last commit February 8, 2026.
How do I install AgentBench?
Clone the repo:
git clone https://github.com/THUDM/AgentBench.gitThen follow the README in the cloned directory.Is AgentBench actively maintained?
stalled — no commits for an extended period
Created
Updated
Last commit
//COMMENTS · 0
Sign in to join the discussion