2 citations · 2 across the 4 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024★ 2 cited
Xmodel-VLM: A Simple Baseline for Multimodal Vision Language Model
Wanting Xu, Yang Liu, Langping He +2
We introduce Xmodel-VLM, a cutting-edge multimodal vision language model. It is designed for efficient deployment on consumer GPU servers. Our work directly confronts a pivotal ind…
cs.CV2024
Event-Based Visual Odometry on Non-Holonomic Ground Vehicles
Wanting Xu, Si'ao Zhang, Li Cui +2
Despite the promise of superior performance under challenging conditions, event-based motion estimation remains a hard problem owing to the difficulty of extracting and tracking st…
cs.CV2024
Tight Fusion of Events and Inertial Measurements for Direct Velocity Estimation
Wanting Xu, Xin Peng, Laurent Kneip
Traditional visual-inertial state estimation targets absolute camera poses and spatial landmark locations while first-order kinematics are typically resolved as an implicitly estim…