activity
20182021
most citedExploring the Best Loss Function for DNN-Based Low-latency Speech Enhancement with Temporal Convolutional Networks

43 citations · 43 across the 2 of their papers we have counts for

collaborators

5 papers

cs.SD2021

Generalized Spoofing Detection Inspired from Audio Generation Artifacts

Yang Gao, Tyler Vuong, Mahsa Elyasi +2

State-of-the-art methods for audio generation suffer from fingerprint artifacts and repeated inconsistencies across temporal and spectral domains. Such artifacts could be well capt…

eess.AS2021

A Modulation-Domain Loss for Neural-Network-based Real-time Speech Enhancement

Tyler Vuong, Yangyang Xia, Richard M. Stern

We describe a modulation-domain loss function for deep-learning-based speech enhancement systems. Learnable spectro-temporal receptive fields (STRFs) were adapted to optimize for a…

eess.AS2020

Learnable Spectro-temporal Receptive Fields for Robust Voice Type Discrimination

Tyler Vuong, Yangyang Xia, Richard Stern

Voice Type Discrimination (VTD) refers to discrimination between regions in a recording where speech was produced by speakers that are physically within proximity of the recording…

eess.AS202043 cited

Exploring the Best Loss Function for DNN-Based Low-latency Speech Enhancement with Temporal Convolutional Networks

Yuichiro Koyama, Tyler Vuong, Stefan Uhlich +1

Recently, deep neural networks (DNNs) have been successfully used for speech enhancement, and DNN-based speech enhancement is becoming an attractive research area. While time-frequ…

cs.CV2018

Natural Language Person Search Using Deep Reinforcement Learning

Ankit Shah, Tyler Vuong

Recent success in deep reinforcement learning is having an agent learn how to play Go and beat the world champion without any prior knowledge of the game. In that task, the agent h…