2 papers
cs.CV2025
VP Lab: a PEFT-Enabled Visual Prompting Laboratory for Semantic Segmentation
Niccolo Avogaro, Thomas Frick, Yagmur G. Cinar +12
Large-scale pretrained vision backbones have transformed computer vision by providing powerful feature extractors that enable various downstream tasks, including training-free appr…
cs.CV2025
Show or Tell? Effectively prompting Vision-Language Models for semantic segmentation
Niccolo Avogaro, Thomas Frick, Mattia Rigotti +5
Large Vision-Language Models (VLMs) are increasingly being regarded as foundation models that can be instructed to solve diverse tasks by prompting, without task-specific training.…