2 papers
cs.CV2025
DiVE: Efficient Multi-View Driving Scenes Generation Based on Video Diffusion Transformer
Junpeng Jiang, Gangyi Hong, Miao Zhang +4
Collecting multi-view driving scenario videos to enhance the performance of 3D visual perception tasks presents significant challenges and incurs substantial costs, making generati…
cs.CV2024
DiVE: DiT-based Video Generation with Enhanced Control
Junpeng Jiang, Gangyi Hong, Lijun Zhou +10
Generating high-fidelity, temporally consistent videos in autonomous driving scenarios faces a significant challenge, e.g. problematic maneuvers in corner cases. Despite recent vid…