1 paper · 1 filter
Rishabh Jain, Naomi Harte
Advances in self-supervised encoders have improved Visual Speech Recognition (VSR). Recent approaches integrating these encoders with LLM decoders improves transcription accuracy;…