2
Publications
58
Citations
2
H-Index
2023
Active since
Publications per year
2023–2023
2
2
AgentBench: Evaluating LLMs as Agents
Xiao Liu, Hao Yu, Hanchen Zhang et al. · arXiv (Cornell University) · 2023 · 40 citations · Full text
SafetyBench: Evaluating the Safety of Large Language Models
Zhexin Zhang, Leqi Lei, Lindong Wu et al. · arXiv (Cornell University) · 2023 · 18 citations · Full text
Rows per page
1–2 of 2