2 papers
cs.CV2026
VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving
Zhefan Xu, Ghassen Jerfel, Marina Haliem +3
The rapid growth of autonomous driving datasets has enabled the scaling of powerful motion forecasting models. While large-scale pretraining provides strong performance, the standa…
cs.LG2024
An End-to-End Reinforcement Learning Based Approach for Micro-View Order-Dispatching in Ride-Hailing
Xinlang Yue, Yiran Liu, Fangzhou Shi +4
Assigning orders to drivers under localized spatiotemporal context (micro-view order-dispatching) is a major task in Didi, as it influences ride-hailing service experience. Existin…