3 papers
cs.CV2026
The DAWN of World-Action Interactive Models
Hongbo Lu, Liang Yao, Chenghao He +6
A plausible scene evolution depends on the maneuver being considered, while a good maneuver depends on how the scene may evolve. Existing World Action Models (WAMs) largely miss th…
cs.CV2026
VisionNVS: Self-Supervised Inpainting for Novel View Synthesis under the Virtual-Shift Paradigm
Hongbo Lu, Liang Yao, Chenghao He +4
A fundamental bottleneck in Novel View Synthesis (NVS) for autonomous driving is the inherent supervision gap on novel trajectories: models are tasked with synthesizing unseen view…
cs.CV2024
SPEAK: Speech-Driven Pose and Emotion-Adjustable Talking Head Generation
Changpeng Cai, Guinan Guo, Jiao Li +7
Most earlier researches on talking face generation have focused on the synchronization of lip motion and speech content. However, head pose and facial emotions are equally importan…