1 paper
Yale Song, Yiwen Song, Nick Losier +13
While diffusion models generate high-fidelity video clips, transforming them into coherent storytelling engines remains challenging. Current agentic pipelines automate this via cha…