Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
NarrativeWorldBench: A Frontier-Saturated Benchmark and a Latent World Model for Long-Horizon Co-Creative Audio Drama
Logan Mann, Abdur Rahman, Mohammad Saifullah +2
Long-form serialized audio drama, with arcs that run for 200 to 800 episodes, is a major creative medium and a setting where frontier large language models (LLMs) fail. We benchmar…
cs.CL2025
Don't Think of the White Bear: Ironic Negation in Transformer Models Under Cognitive Load
Logan Mann, Nayan Saxena, Sarah Tandon +3
Negation instructions such as 'do not mention ' can paradoxically increase the accessibility of in human thought, a phenomenon known as ironic rebound. Large language models…