3 papers
cs.RO2026
ReflectDrive-2: Reinforcement-Learning-Aligned Self-Editing for Discrete Diffusion Driving
Huimin Wang, Yue Wang, Bihao Cui +7
We introduce ReflectDrive-2, a masked discrete diffusion planner with separate action expert for autonomous driving that represents plans as discrete trajectory tokens and generate…
cs.CV2026
AgentMV: A State-Guided Multi-Agent Framework for Budget-Aware Music Video Generation
Huimin Wang, Leilei Ouyang, Chang Xia +3
Generating a complete music video from a song requires more than synthesizing visually plausible clips for individual lyric prompts. A practical system must maintain long-range vis…
cs.RO2025
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
Pengxiang Li, Yinan Zheng, Yue Wang +6
End-to-End (E2E) solutions have emerged as a mainstream approach for autonomous driving systems, with Vision-Language-Action (VLA) models representing a new paradigm that leverages…