67 citations · 139 across the 28 of their papers we have counts for
3 papers · 1 filter
End-to-end spoofing detection with raw waveform CLDNNs
Heinrich Dinkel, Nanxin Chen, Yanmin Qian +1
Albeit recent progress in speaker verification generates powerful models, malicious attacks in the form of spoofed speech, are generally not coped with. Recent results in ASVSpoof2…
Multiple Sound Sources Localization from Coarse to Fine
Rui Qian, Di Hu, Heinrich Dinkel +3
How to visually localize multiple sound sources in unconstrained videos is a formidable problem, especially when lack of the pairwise sound-object annotations. To solve this proble…
Voice activity detection in the wild via weakly supervised sound event detection
Heinrich Dinkel, Yefei Chen, Mengyue Wu +1
Traditional supervised voice activity detection (VAD) methods work well in clean and controlled scenarios, with performance severely degrading in real-world applications. One possi…