10 papers
Harness the Memory: A Holistic Evaluation of Memory Substrates in Memory Agents
Wei-Chieh Huang, Weizhi Zhang, Yuchen Wu +12
Memory is becoming core infrastructure for long-horizon LLM agents, yet existing evaluations offer limited guidance on which memory substrate, namely the underlying medium in which…
QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents
Ye Yuan, Rui Song, Weien Li +12
Social deduction games have become a popular testbed for probing reasoning, deception, coordination, and belief modeling in Large Language Model (LLM) agents. However, most environ…
Beyond Message Passing: A Semantic View of Agent Communication Protocols
Dun Yuan, Fuyuan Lyu, Ye Yuan +11
Agent communication protocols are becoming critical infrastructure for large language model (LLM) systems that must use tools, coordinate with other agents, and operate across hete…
"I Use ChatGPT to Humanize My Words": Affordances and Risks of ChatGPT to Autistic Users
Renkai Ma, Ben Zefeng Zhang, Chen Chen +4
Large Language Model (LLM) chatbots like ChatGPT have emerged as cognitive scaffolding for autistic users, yet the tension between their utility and risk remains under-articulated.…
LLM Use for Mental Health: Crowdsourcing Users' Sentiment-based Perspectives and Values from Social Discussions
Lingyao Li, Xiaoshan Huang, Renkai Ma +4
Large language models (LLMs) chatbots like ChatGPT are increasingly used for mental health support. They offer accessible, therapeutic support but also raise concerns about misinfo…
Can LLM Agents Really Debate? A Controlled Study of Multi-Agent Debate in Logical Reasoning
Haolun Wu, Zhenkun Li, Lingyao Li
Multi-agent debate (MAD) has recently emerged as a promising framework for improving the reasoning performance of large language models (LLMs). Yet, whether LLM agents can genuinel…