1 paper · 1 filter
Ziqin Huang, Yingyue Li, Chenyangguang Zhang +6
Intermediate representations are key to bridging the modality gap between generalizable manipulation policies and large-scale pretrained vision-language models (VLMs). Among these,…