3 citations · 3 across the 6 of their papers we have counts for
1 paper · 1 filter
Hao Cheng, Erjia Xiao, Yichi Wang +12
Recently, driven by advancements in Multimodal Large Language Models (MLLMs), Vision Language Action Models (VLAMs) are being proposed to achieve better performance in open-vocabul…