4 papers
ClustViT: Clustering-based Token Merging for Semantic Segmentation
Fabio Montello, Ronja Güldenring, Lazaros Nalpantidis
Vision Transformers can achieve high accuracy and strong generalization across various contexts, but their practical applicability on real-world robotic systems is limited due to t…
A Survey on Dynamic Neural Networks: from Computer Vision to Multi-modal Sensor Fusion
Fabio Montello, Ronja Güldenring, Simone Scardapane +1
Model compression is essential in the deployment of large Computer Vision models on embedded devices. However, static optimization techniques (e.g. pruning, quantization, etc.) neg…
Spiking Patches: Asynchronous, Sparse, and Efficient Tokens for Event Cameras
Christoffer Koo Ãhrstrøm, Ronja Güldenring, Lazaros Nalpantidis
We propose tokenization of events and present a tokenizer, Spiking Patches, specifically designed for event cameras. Given a stream of asynchronous and spatially sparse events, our…
From Web Data to Real Fields: Low-Cost Unsupervised Domain Adaptation for Agricultural Robots
Vasileios Tzouras, Lazaros Nalpantidis, Ronja Güldenring
In precision agriculture, vision models often struggle with new, unseen fields where crops and weeds have been influenced by external factors, resulting in compositions and appeara…