From the 1 of 4 linked papers with an AI index.
4 papers
NormAct: Benchmarking Embodied Agents' Proactive Compliance with Unspoken Social Norms
Shiyun Zhao, Xinwei Song, Tianyu Guo +7
The paper presents NormAct, a benchmark for evaluating whether embodied planners using multimodal large language models can infer and follow hidden social norms while completing ta…
Credibility Governance: A Social Mechanism for Collective Self-Correction under Weak Truth Signals
Wanying He, Yanxi Lin, Ziheng Zhou +5
Online platforms increasingly rely on opinion aggregation to allocate real-world attention and resources, yet common signals such as engagement votes or capital-weighted commitment…
Evaluating Generalization Capabilities of LLM-Based Agents in Mixed-Motive Scenarios Using Concordia
Chandler Smith, Marwa Abdulhai, Manfred Diaz +83
Large Language Model (LLM) agents have demonstrated impressive capabilities for social interaction and are increasingly being deployed in situations where they might engage with bo…
Are the Values of LLMs Structurally Aligned with Humans? A Causal Perspective
Yipeng Kang, Junqi Wang, Yexin Li +8
As large language models (LLMs) become increasingly integrated into critical applications, aligning their behavior with human values presents significant challenges. Current method…