1 paper
Anurag Das, Xinting Hu, Li Jiang +1
Recent approaches have shown that large-scale vision-language models such as CLIP can improve semantic segmentation performance. These methods typically aim for pixel-level vision-…