1 paper
Nicolas Harvey Chapman, Feras Dayoub, Will Browne +1
A domain shift exists between the large-scale, internet data used to train a Vision-Language Model (VLM) and the raw image streams collected by a robot. Existing adaptation strateg…