1 citations · 1 across the 3 of their papers we have counts for
3 papers
A Real-Time Active Speaker Detection System Integrating an Audio-Visual Signal with a Spatial Querying Mechanism
Ilya Gurvich, Ido Leichter, Dharmendar Reddy Palle +6
We introduce a distinctive real-time, causal, neural network-based active speaker detection system optimized for low-power edge computing. This system drives a virtual cinematograp…
LSTM-based Video Quality Prediction Accounting for Temporal Distortions in Videoconferencing Calls
Gabriel Mittag, Babak Naderi, Vishak Gopal +1
Current state-of-the-art video quality models, such as VMAF, give excellent prediction results by comparing the degraded video with its reference video. However, they do not consid…
ICASSP 2023 Deep Noise Suppression Challenge
Harishchandra Dubey, Ashkan Aazami, Vishak Gopal +9
Deep Speech Enhancement Challenge is the 5th edition of deep noise suppression (DNS) challenges organized at ICASSP 2023 Signal Processing Grand Challenges. DNS challenges were org…