1 paper · 1 filter
Mohamed L. Mekhalfi, Mohamad M. Al Rahhal, Yakoub Bazi +4
Vision-language models like CLIP have shown sig- nificant potential in handling natural images, yet their perfor- mance is often limited by the distinct characteristics of satellit…