55 citations · 98 across the 2 of their papers we have counts for
3 papers
cs.AI2023★ 55 cited
AgentBench: Evaluating LLMs as Agents
Xiao Liu, Hao Yu, Hanchen Zhang +19
The potential of Large Language Model (LLM) as agents has been widely acknowledged recently. Thus, there is an urgent need to quantitatively \textit{evaluate LLMs as agents} on cha…
cs.CY2023★ 43 cited
Deceptive AI Ecosystems: The Case of ChatGPT
Xiao Zhan, Yifan Xu, Stefan Sarkadi
ChatGPT, an AI chatbot, has gained popularity for its capability in generating human-like responses. However, this feature carries several risks, most notably due to its deceptive…
cs.HC2020
Improving Workflow Integration with xPath: Design and Evaluation of a Human-AI Diagnosis System in Pathology
Hongyan Gu, Yuan Liang, Yifan Xu +12
Recent developments in AI have provided assisting tools to support pathologists' diagnoses. However, it remains challenging to incorporate such tools into pathologists' practice; o…