2 citations · 2 across the 1 of their papers we have counts for
1 paper
Hao Bai, Yifei Zhou, Mert Cemri +4
Training corpuses for vision language models (VLMs) typically lack sufficient amounts of decision-centric data. This renders off-the-shelf VLMs sub-optimal for decision-making task…