1 paper · 1 filter
Siddharth Karamcheti, Suraj Nair, Ashwin Balakrishna +3
Visually-conditioned language models (VLMs) have seen growing adoption in applications such as visual dialogue, scene understanding, and robotic task planning; adoption that has fu…