3 citations · 3 across the 3 of their papers we have counts for
1 paper · 1 filter
Keyang Xuan, Pengda Wang, Chongrui Ye +3
Large language models (LLMs) are increasingly evaluated in interactive environments to test their social intelligence. However, existing benchmarks often assume idealized communica…