3 papers
cs.LG2025
SafeArena: Evaluating the Safety of Autonomous Web Agents
Ada Defne Tur, Nicholas Meade, Xing Han Lù +6
LLM-based agents are becoming increasingly proficient at solving web-based tasks. With this capability comes a greater risk of misuse for malicious purposes, such as posting misinf…
cs.CL2025
Language Models Largely Exhibit Human-like Constituent Ordering Preferences
Ada Defne Tur, Gaurav Kamath, Siva Reddy
Though English sentences are typically inflexible vis-à-vis word order, constituents often show far more variability in ordering. One prominent theory presents the notion that cons…
cs.CL2024
ProGRes: Prompted Generative Rescoring on ASR n-Best
Ada Defne Tur, Adel Moumen, Mirco Ravanelli
Large Language Models (LLMs) have shown their ability to improve the performance of speech recognizers by effectively rescoring the n-best hypotheses generated during the beam sear…