Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Conformal Cross-Modal Active Learning
Huy Hoang Nguyen, Cédric Jung, Shirin Salehi +3
Foundation models for vision have transformed visual recognition with powerful pretrained representations and strong zero-shot capabilities, yet their potential for data-efficient…
cs.CV2024
Enhancing CTC-Based Visual Speech Recognition
Hendrik Laux, Anke Schmeink
This paper presents LiteVSR2, an enhanced version of our previously introduced efficient approach to Visual Speech Recognition (VSR). Building upon our knowledge distillation frame…