2 papers
cs.CV2025
Gestura: A LVLM-Powered System Bridging Motion and Semantics for Real-Time Free-Form Gesture Understanding
Zhuoming Li, Aitong Liu, Mengxi Jia +5
Free-form gesture understanding is highly appealing for human-computer interaction, as it liberates users from the constraints of predefined gesture categories. However, the sole e…
cs.LG2025
ECVL-ROUTER: Scenario-Aware Routing for Vision-Language Models
Xin Tang, Youfang Han, Fangfei Gou +7
Vision-Language Models (VLMs) excel in diverse multimodal tasks. However, user requirements vary across scenarios, which can be categorized into fast response, high-quality output,…