1 paper · 1 filter
Ruoshi Liu, Alper Canberk, Shuran Song +1
Vision foundation models trained on massive amounts of visual data have shown unprecedented reasoning and planning skills in open-world settings. A key challenge in applying them t…