43 citations · 43 across the 2 of their papers we have counts for
5 papers
Generalized Spoofing Detection Inspired from Audio Generation Artifacts
Yang Gao, Tyler Vuong, Mahsa Elyasi +2
State-of-the-art methods for audio generation suffer from fingerprint artifacts and repeated inconsistencies across temporal and spectral domains. Such artifacts could be well capt…
A Modulation-Domain Loss for Neural-Network-based Real-time Speech Enhancement
Tyler Vuong, Yangyang Xia, Richard M. Stern
We describe a modulation-domain loss function for deep-learning-based speech enhancement systems. Learnable spectro-temporal receptive fields (STRFs) were adapted to optimize for a…
Learnable Spectro-temporal Receptive Fields for Robust Voice Type Discrimination
Tyler Vuong, Yangyang Xia, Richard Stern
Voice Type Discrimination (VTD) refers to discrimination between regions in a recording where speech was produced by speakers that are physically within proximity of the recording…
Exploring the Best Loss Function for DNN-Based Low-latency Speech Enhancement with Temporal Convolutional Networks
Yuichiro Koyama, Tyler Vuong, Stefan Uhlich +1
Recently, deep neural networks (DNNs) have been successfully used for speech enhancement, and DNN-based speech enhancement is becoming an attractive research area. While time-frequ…
Natural Language Person Search Using Deep Reinforcement Learning
Ankit Shah, Tyler Vuong
Recent success in deep reinforcement learning is having an agent learn how to play Go and beat the world champion without any prior knowledge of the game. In that task, the agent h…