activity
20242026
collaborators

7 papers

cs.CV2026

Thermo-VL: Extending Vision-Language Models to Thermal Infrared Perception

Rusiru Thushara, Yasiru Ranasinghe, Jay Paranjape +1

Vision-language models (VLMs) often fail under low illumination because their visual grounding is learned predominantly from RGB imagery, whereas thermal infrared preserves complem…

eess.IV2026

Next-Acceleration-Scale Prediction for Autoregressive MRI Reconstruction

Yilmaz Korkmaz, Vishal M. Patel

MRI reconstruction is an inherently ill-posed inverse problem, since incomplete measurements admit many plausible solutions. This ambiguity becomes more severe under high accelerat…

cs.CV2025

UniRes: Universal Image Restoration for Complex Degradations

Mo Zhou, Keren Ye, Mauricio Delbracio +3

Real-world image restoration is hampered by diverse degradations stemming from varying capture conditions, capture devices and post-processing pipelines. Existing works make improv…

cs.CV2025

Reference-Guided Identity Preserving Face Restoration

Mo Zhou, Keren Ye, Viraj Shah +5

Preserving face identity is a critical yet persistent challenge in diffusion-based image restoration. While reference faces offer a path forward, existing reference-based methods o…

cs.CV2025

Latent Feature-Guided Diffusion Models for Shadow Removal

Kangfu Mei, Luis Figueroa, Zhe Lin +3

Recovering textures under shadows has remained a challenging problem due to the difficulty of inferring shadow-free scenes from shadow images. In this paper, we propose the use of…

cs.CV2025

The Power of Context: How Multimodality Improves Image Super-Resolution

Kangfu Mei, Hossein Talebi, Mojtaba Ardakani +3

Single-image super-resolution (SISR) remains challenging due to the inherent difficulty of recovering fine-grained details and preserving perceptual quality from low-resolution inp…