6 papers · 1 filter
Beyond Natural-Image Foundation Models: Benchmarking Satellite Pretraining for Ophthalmic Image Analysis
Lovre Antonio Budimir, Mingya Alexa Gong, Alyssa Foong Quinney +5
Vision Foundation Models (VFMs) have emerged as a promising approach in medical imaging, producing broadly applicable systems that can be efficiently adapted across diverse imaging…
Representation Transfer of Foundation Models for Ultra-Widefield Retinal Imaging
Mingya Alexa Gong, Da Ma, Lovre Antonio Budimir +7
Despite the widespread adoption of foundation models as feature extractors for medical imaging, relatively little is understood about how different pretraining strategies influence…
Sparse Code Uplifting for Efficient 3D Language Gaussian Splatting
Lovre Antonio Budimir, Yushi Guan, Steve Ryhner +2
3D Language Gaussian Splatting (3DLGS) augments 3D Gaussian Splatting with language-aligned visual features for open-vocabulary 3D scene understanding. A core challenge is efficien…
OOS-DSD: Improving Out-of-stock Detection in Retail Images using Auxiliary Tasks
Franko Šikić, Sven Lončarić
Out-of-stock (OOS) detection is a very important retail verification process that aims to infer the unavailability of products in their designated areas on the shelf. In this paper…
DHECA-SuperGaze: Dual Head-Eye Cross-Attention and Super-Resolution for Unconstrained Gaze Estimation
Franko Šikić, Donik Vršnak, Sven Lončarić
Unconstrained gaze estimation is the process of determining where a subject is directing their visual attention in uncontrolled environments. Gaze estimation systems are important…
A Survey on Deep Learning-based Gaze Direction Regression: Searching for the State-of-the-art
Franko Šikić, Donik Vršnak, Sven Lončarić
In this paper, we present a survey of deep learning-based methods for the regression of gaze direction vector from head and eye images. We describe in detail numerous published met…