2 papers
cs.RO2025
DiffE2E: Rethinking End-to-End Driving with a Hybrid Action Diffusion and Supervised Policy
Rui Zhao, Yuze Fan, Ziguo Chen +2
End-to-end learning has emerged as a transformative paradigm in autonomous driving. However, the inherently multimodal nature of driving behaviors and the generalization challenges…
cs.CV2025
Re-Aligning Language to Visual Objects with an Agentic Workflow
Yuming Chen, Jiangyan Feng, Haodong Zhang +6
Language-based object detection (LOD) aims to align visual objects with language expressions. A large amount of paired data is utilized to improve LOD model generalizations. During…