3 papers
cs.CV2025
Differentiable Hierarchical Visual Tokenization
Marius Aasan, Martine Hjelkrem-Tan, Nico Catalano +2
Vision Transformers rely on fixed patch tokens that ignore the spatial and semantic structure of images. In this work, we introduce an end-to-end differentiable tokenizer that adap…
cs.CV2025
MARS: a Multimodal Alignment and Ranking System for Few-Shot Segmentation
Nico Catalano, Stefano Samele, Paolo Pertino +1
Few Shot Segmentation aims to segment novel object classes given only a handful of labeled examples, enabling rapid adaptation with minimal supervision. Current literature cruciall…
cs.CV2024
More than the Sum of Its Parts: Ensembling Backbone Networks for Few-Shot Segmentation
Nico Catalano, Alessandro Maranelli, Agnese Chiatti +1
Semantic segmentation is a key prerequisite to robust image understanding for applications in \acrlong{ai} and Robotics. \acrlong{fss}, in particular, concerns the extension and op…