35 papers
REFRAMED: Towards Realistic Audio Description Generation for Movies
Igor Sterner, Mirella Lapata, Alex Lascarides +1
Audio Description (AD) is a verbal narration of key visual content in videos, enabling access for visually impaired audiences. Unlike standard video captioning, AD is a structured…
When is Routing Meaningful? Diversity and Robustness in Language Model Societies
Fantine Huot, Michael Kaisers, Mirella Lapata
Routing policies for multi-model systems are evaluated almost exclusively on task accuracy and inference cost. We argue that two properties, orthogonal to performance, determine wh…
Storyline Trees: Hierarchical Representations for Long-Form Narratives
Litu Ou, Mirella Lapata
Long-form narratives are challenging for long-context models because their structure is implicit: events, characters, and plotlines interact across hundreds of pages without the ex…
Lightweight Latent Reasoning for Narrative Tasks
Alexander Gurung, Esmeralda S. Whitammer, Mirella Lapata
Large language models (LLMs) tackle complex tasks by generating long chains of thought or "reasoning traces" that act as latent variables in the generation of an output given a que…
GraphLit: Learning Text-Enriched Dynamic Character Network Representations for Literary Study
Gaspard Michel, Elena V. Epure, Romain Hennequin +2
Methods to represent literary texts as graphs or sequences of graphs mainly focus on representing character interactions, and often overlook another crucial aspect: the textual con…
Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning
Miao Li, Irina Saparina, Alexander Gurung +1
Recent large language models support inputs of up to 10 million tokens, yet they perform poorly on long-context tasks that require complex reasoning. Such tasks can be solved using…