2 papers
cs.CV2026
EG-VQA: Benchmarking Verifiable Video Question Answering with Grounded Temporal Evidence
Linpeng Huang, Weixing Chen, Zexin Chen +2
Recent advances in Video Large Language Models (Video-LLMs) have yielded promising performance on video question answering (VideoQA). Nevertheless, existing benchmarks are predomin…
cs.AI2025
STORY2GAME: Generating (Almost) Everything in an Interactive Fiction Game
Eric Zhou, Shreyas Basavatia, Moontashir Siam +2
We introduce STORY2GAME, a novel approach to using Large Language Models to generate text-based interactive fiction games that starts by generating a story, populates the world, an…