2 papers
cs.CV2026
ScenarioControl: Vision-Language Controllable Vectorized Latent Scenario Generation
Lili Gao, Yanbo Xu, William Koch +8
We introduce ScenarioControl, the first vision-language control mechanism for learned driving scenario generation. Given a text prompt or an input image, Scenario-Control synthesiz…
cs.CV2026
ChopGrad: Pixel-Wise Losses for Latent Video Diffusion via Truncated Backpropagation
Dmitriy Rivkin, Parker Ewen, Lili Gao +5
Recent video diffusion models achieve high-quality generation through recurrent frame processing where each frame generation depends on previous frames. However, this recurrent mec…