4 papers
Would you still call this Dax? Novel Visual References in VLMs and Humans
Ada Defne Tür, Gaurav Kamath, Joyce Chai +2
Vision-language models (VLMs), like human learners, are frequently exposed to new visual concepts, but how they map novel visual references to language after exposure remains large…
SafeArena: Evaluating the Safety of Autonomous Web Agents
Ada Defne Tur, Nicholas Meade, Xing Han Lù +6
LLM-based agents are becoming increasingly proficient at solving web-based tasks. With this capability comes a greater risk of misuse for malicious purposes, such as posting misinf…
Language Models Largely Exhibit Human-like Constituent Ordering Preferences
Ada Defne Tur, Gaurav Kamath, Siva Reddy
Though English sentences are typically inflexible vis-Ã -vis word order, constituents often show far more variability in ordering. One prominent theory presents the notion that con…
ProGRes: Prompted Generative Rescoring on ASR n-Best
Ada Defne Tur, Adel Moumen, Mirco Ravanelli
Large Language Models (LLMs) have shown their ability to improve the performance of speech recognizers by effectively rescoring the n-best hypotheses generated during the beam sear…