3 papers
cs.CL2025
Are Multimodal Large Language Models Pragmatically Competent Listeners in Simple Reference Resolution Tasks?
Simeon Junker, Manar Ali, Larissa Koch +2
We investigate the linguistic abilities of multimodal large language models in reference resolution tasks featuring simple yet abstract visual stimuli, such as color patches and co…
cs.CL2025
SceneGram: Conceptualizing and Describing Tangrams in Scene Context
Simeon Junker, Sina ZarrieÃ
Research on reference and naming suggests that humans can come up with very different ways of conceptualizing and referring to the same object, e.g. the same abstract tangram shape…
cs.CL2024
Resilience through Scene Context in Visual Referring Expression Generation
Simeon Junker, Sina ZarrieÃ
Scene context is well known to facilitate humans' perception of visible objects. In this paper, we investigate the role of context in Referring Expression Generation (REG) for obje…