24 citations · 51 across the 7 of their papers we have counts for
9 papers · 1 filter
The impact of removing head movements on audio-visual speech enhancement
Zhiqi Kang, Mostafa Sadeghi, Radu Horaud +3
This paper investigates the impact of head movements on audio-visual speech enhancement (AVSE). Although being a common conversational feature, head movements have been ignored by…
Do sound event representations generalize to other audio tasks? A case study in audio transfer learning
Anurag Kumar, Yun Wang, Vamsi Krishna Ithapu +1
Transfer learning is critical for efficient information transfer across multiple related learning problems. A simple, yet effective transfer learning approach utilizes deep neural…
A Sequential Self Teaching Approach for Improving Generalization in Sound Event Recognition
Anurag Kumar, Vamsi Krishna Ithapu
An important problem in machine auditory perception is to recognize and detect sound events. In this paper, we propose a sequential self-teaching approach to learning sounds. Our m…
Secost: Sequential co-supervision for large scale weakly labeled audio event detection
Anurag Kumar, Vamsi Krishna Ithapu
Weakly supervised learning algorithms are critical for scaling audio event detection to several hundreds of sound categories. Such learning models should not only disambiguate soun…
Learning Sound Events From Webly Labeled Data
Anurag Kumar, Ankit Shah, Bhiksha Raj +1
In the last couple of years, weakly labeled learning has turned out to be an exciting approach for audio event detection. In this work, we introduce webly labeled learning for soun…
A Closer Look at Weak Label Learning for Audio Events
Ankit Shah, Anurag Kumar, Alexander G. Hauptmann +1
Audio content analysis in terms of sound events is an important research problem for a variety of applications. Recently, the development of weak labeling approaches for audio or s…