benchmark 1contextual cue generation 1embodied planning 1multimodal language models 1social norm compliance 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
NormAct: Benchmarking Embodied Agents' Proactive Compliance with Unspoken Social Norms
Shiyun Zhao, Xinwei Song, Tianyu Guo +7
The paper presents NormAct, a benchmark for evaluating whether embodied planners using multimodal large language models can infer and follow hidden social norms while completing ta…
cs.AI2025
TongSIM: A General Platform for Simulating Intelligent Machines
Zhe Sun, Kunlun Wu, Chuanjian Fu +24
As artificial intelligence (AI) rapidly advances, especially in multimodal large language models (MLLMs), research focus is shifting from single-modality text processing to the mor…