1 paper
Kodai Kawamura, Yuta Goto, Rintaro Yanagi +2
Pre-trained Vision-Language Models (VLMs) exhibit strong generalization capabilities, enabling them to recognize a wide range of objects across diverse domains without additional t…