Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Beyond Coherence: Benchmarking Professional Editing-Technique Execution in Multi-Shot Audio-Video Generation
Tianyi Zeng, Junchao Liao, Yujie Wei +9
Recent multi-shot audio-video generators can produce increasingly coherent and cinematic outputs, but coherence does not imply the ability to execute editing techniques. Profession…
cs.AI2026
World of Workflows: A Benchmark for Bringing World Models to Enterprise Systems
Lakshya Gupta, Litao Li, Yizhe Liu +5
Frontier large language models (LLMs) excel as autonomous agents in many domains, yet they remain untested in complex enterprise systems where hidden workflows create cascading eff…