Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Investigating Multimodal Informativity under Different Partner Visibility Conditions in Video-Mediated Dialogue
Esam Ghaleb, Hugh Mee Wong, Kristina Kobrock
Situated language use is multimodal and embodied. For example, gestures can carry information that is absent or underspecified in the speech signal, yet dialogue models typically r…
cs.CL2026
Aligned but Not Partner-Specific: Distinguishing How Multimodal LLM Agents Succeed in Reference Games Without Human-Like Conventions
Po-Ya Angela Wang, Chinmaya Mishra, Aslı Ãzyürek +2
Repeated reference games test whether interlocutors replace their initially long descriptions with shorter, partner-specific conventions grounded in shared interaction history. Pri…
cs.CL2025
LLMs instead of Human Judges? A Large Scale Empirical Study across 20 NLP Evaluation Tasks
Anna Bavaresco, Raffaella Bernardi, Leonardo Bertolazzi +17
There is an increasing trend towards evaluating NLP models with LLMs instead of human judgments, raising questions about the validity of these evaluations, as well as their reprodu…