3 papers
cs.AI2026
Narrative World Model: Narratology-Grounded Writer Memory for Long-Form Fiction
Mohammad Saifullah, Thomas Kornmaier, Taaha Kazi +3
Long-form fiction writers need memory that answers multi-hop questions about evolving story state: who knows a secret and when they learned it, whether an event preceded the narrat…
cs.CL2026
NarrativeWorldBench: A Frontier-Saturated Benchmark and a Latent World Model for Long-Horizon Co-Creative Audio Drama
Logan Mann, Abdur Rahman, Mohammad Saifullah +2
Long-form serialized audio drama, with arcs that run for 200 to 800 episodes, is a major creative medium and a setting where frontier large language models (LLMs) fail. We benchmar…
cs.CL2024
Large Language Models as User-Agents for Evaluating Task-Oriented-Dialogue Systems
Taaha Kazi, Ruiliang Lyu, Sizhe Zhou +2
Traditionally, offline datasets have been used to evaluate task-oriented dialogue (TOD) models. These datasets lack context awareness, making them suboptimal benchmarks for convers…