1 paper
Ye Won Byun, Cathy Jiao, Shahriar Noroozizadeh +2
We introduce a simple method that employs pre-trained CLIP encoders to enhance model generalization in the ALFRED task. In contrast to previous literature where CLIP replaces the v…