The Agent Index / Agent Frameworks / #258
THUDM/AgentBench
by THUDM · Agent Frameworks · updated 7mo ago
A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)
37
momentum
3,724
stars
279
forks
#258
rank
chatgptgpt-4llmllm-agent
View on GitHub →