most citedThe PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies

1 citations · 1 across the 7 of their papers we have counts for

collaborators

18 papers

cs.CY2026

When Should AI Read the Room? Public Perceptions of Social Intelligence in AI Agents

Leena Mathur, Jenny T. Liang, Vasudha Varadarajan +6

AI researchers have been advancing socially intelligent AI agents (Social-AI) across embodiments, from chatbots to physical robots. As Social-AI is increasingly deployed in everyda…

cs.MA2026

SOTOPIA-TOM: Evaluating Information Management in Multi-Agent Interaction with Theory of Mind

Yashwanth YS, Ruichen Wang, Shihua Zeng +4

As LLM-based agents are increasingly interacting in multi-party settings, they need to properly handle information asymmetry, i.e., knowing when and to whom to disclose information…

cs.CL2026

On Safety Risks in Experience-Driven Self-Evolving Agents

Weixiang Zhao, Yichen Zhang, Yingshuo Wang +8

Experience-driven self-evolution has emerged as a promising paradigm for improving the autonomy of large language model agents, yet its reliance on self-curated experience introduc…

cs.CL2026

Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies

Myke C. Cohen, Mingqian Zheng, Neel Bhandari +6

AI design characteristics and human personality traits each impact the quality and outcomes of human-AI interactions. However, their relative and joint impacts are underexplored in…

cs.AI2026

GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses

Jimin Mun, Chani Jung, Xuhui Zhou +2

While LLMs hold significant potential to transform scientific research, we advocate for their use to augment and empower researchers rather than to automate research without human…

cs.CL20261 cited

The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies

Jiaxu Zhou, Jen-tse Huang, Xuhui Zhou +5

Large language models (LLMs) are increasingly deployed to simulate human collective behaviors, yet the methodological rigor of these "AI societies" remains under-explored. Through…