1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2024
Unveiling Uncertainty: A Deep Dive into Calibration and Performance of Multimodal Large Language Models
Zijun Chen, Wenbo Hu, Guande He +3
Multimodal large language models (MLLMs) combine visual and textual data for tasks such as image captioning and visual question answering. Proper uncertainty calibration is crucial…
cs.CV2024★ 1 cited
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving
Xiaosong Jia, Shaoshuai Shi, Zijun Chen +4
As an essential task in autonomous driving (AD), motion prediction aims to predict the future states of surround objects for navigation. One natural solution is to estimate the pos…