3 papers
cs.CL2025
Length Aware Speech Translation for Video Dubbing
Harveen Singh Chadha, Aswin Shanmugam Subramanian, Vikas Joshi +4
In video dubbing, aligning translated audio with the source audio is a significant challenge. Our focus is on achieving this efficiently, tailored for real-time, on-device video du…
eess.AS2025
Building English ASR model with regional language support
Purvi Agrawal, Vikas Joshi, Bharati Patidar +2
In this paper, we present a novel approach to developing an English Automatic Speech Recognition (ASR) system that can effectively handle Hindi queries, without compromising its pe…
cs.CL2025
Addressing speaker gender bias in large scale speech translation systems
Shubham Bansal, Vikas Joshi, Harveen Chadha +2
This study addresses the issue of speaker gender bias in Speech Translation (ST) systems, which can lead to offensive and inaccurate translations. The masculine bias often found in…