5 papers
Can LoRA Fusion Support Cross-Domain Tasks in Cloud-Edge Collaboration?
Yatong Wang, Fali Wang, Naibin Gu +6
Cloud-hosted large language models (LLMs) commonly rely on LoRA for domain adaptation, yet domain data are distributed across multiple edge devices and cannot be uploaded due to pr…
MathSE: Improving Multimodal Mathematical Reasoning via Self-Evolving Iterative Reflection and Reward-Guided Fine-Tuning
Jinhao Chen, Zhen Yang, Jianxin Shi +2
Multimodal large language models (MLLMs) have demonstrated remarkable capabilities in vision-language answering tasks. Despite their strengths, these models often encounter challen…
DECAMP: Towards Scene-Consistent Multi-Agent Motion Prediction with Disentangled Context-Aware Pre-Training
Jianxin Shi, Zengqi Peng, Xiaolong Chen +2
Trajectory prediction is a critical component of autonomous driving, essential for ensuring both safety and efficiency on the road. However, traditional approaches often struggle w…
MonoGlass3D: Monocular 3D Glass Detection with Plane Regression and Adaptive Feature Fusion
Kai Zhang, Guoyang Zhao, Jianxing Shi +3
Detecting and localizing glass in 3D environments poses significant challenges for visual perception systems, as the optical properties of glass often hinder conventional sensors f…
Motion Forecasting for Autonomous Vehicles: A Survey
Jianxin Shi, Jinhao Chen, Yuandong Wang +4
In recent years, the field of autonomous driving has attracted increasingly significant public interest. Accurately forecasting the future behavior of various traffic participants…