2 papers
cs.CV2026
ClustViT: Clustering-based Token Merging for Semantic Segmentation
Fabio Montello, Ronja Güldenring, Lazaros Nalpantidis
Vision Transformers can achieve high accuracy and strong generalization across various contexts, but their practical applicability on real-world robotic systems is limited due to t…
cs.CV2026
A Survey on Dynamic Neural Networks: from Computer Vision to Multi-modal Sensor Fusion
Fabio Montello, Ronja Güldenring, Simone Scardapane +1
Model compression is essential in the deployment of large Computer Vision models on embedded devices. However, static optimization techniques (e.g. pruning, quantization, etc.) neg…