1 citations · 1 across the 4 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
QuoVLA: Quotient Space for Vision-Language-Action Models
Xuan Wang, Yinan Wu, Haoran Duan +1
Vision-Language-Action (VLA) models commonly adapt pretrained Vision-Language Models (VLMs) to robot control by mapping visual observations and language instructions to continuous…
cs.CV2025★ 1 cited
Enhancing Target-unspecific Tasks through a Features Matrix
Fangming Cui, Yonggang Zhang, Xuan Wang +2
Recent developments in prompt learning of large Vision-Language Models (VLMs) have significantly improved performance in target-specific tasks. However, these prompting methods oft…
cs.CV2025
Generalizable Prompt Learning of CLIP: A Brief Overview
Fangming Cui, Yonggang Zhang, Xuan Wang +2
Existing vision-language models (VLMs) such as CLIP have showcased an impressive capability to generalize well across various downstream tasks. These models leverage the synergy be…