5 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CV2025
TokenFLEX: Unified VLM Training for Flexible Visual Tokens Inference
Junshan Hu, Jialiang Mao, Zhikang Liu +3
Conventional Vision-Language Models(VLMs) typically utilize a fixed number of vision tokens, regardless of task complexity. This one-size-fits-all strategy introduces notable ineff…
cs.RO2024★ 5 cited
PlanAgent: A Multi-modal Large Language Agent for Closed-loop Vehicle Motion Planning
Yupeng Zheng, Zebin Xing, Qichao Zhang +8
Vehicle motion planning is an essential component of autonomous driving technology. Current rule-based vehicle motion planning methods perform satisfactorily in common scenarios bu…
cs.RO2023★ 2 cited
Planning-inspired Hierarchical Trajectory Prediction for Autonomous Driving
Ding Li, Qichao Zhang, Zhongpu Xia +4
Recently, anchor-based trajectory prediction methods have shown promising performance, which directly selects a final set of anchors as future intents in the spatio-temporal couple…