Showing 2025Show all
3 papers · 1 filter
cs.CL2025
Lightweight Latent Reasoning for Narrative Tasks
Alexander Gurung, Esmeralda S. Whitammer, Mirella Lapata
Large language models (LLMs) tackle complex tasks by generating long chains of thought or "reasoning traces" that act as latent variables in the generation of an output given a que…
cs.AI2025
SciTrek: Evaluating and Improving Long-Context Numerical Reasoning over Scientific Articles
Miao Li, Alexander Gurung, Irina Saparina +1
We introduce SciTrek, a synthetic question-answering dataset for assessing and improving long-context numerical reasoning in large language models (LLMs). Existing long-context dat…
cs.CL2025
Learning to Reason for Long-Form Story Generation
Alexander Gurung, Mirella Lapata
Generating high-quality stories spanning thousands of tokens requires competency across a variety of skills, from tracking plot and character arcs to keeping a consistent and engag…