514 citations · 674 across the 27 of their papers we have counts for
4 papers · 2 filters
CITISEN: A Deep Learning-Based Speech Signal-Processing Mobile Application
Yu-Wen Chen, Kuo-Hsuan Hung, You-Jin Li +6
This study presents a deep learning-based speech signal-processing mobile application known as CITISEN. The CITISEN provides three functions: speech enhancement (SE), model adaptat…
Waveform-based Voice Activity Detection Exploiting Fully Convolutional networks with Multi-Branched Encoders
Cheng Yu, Kuo-Hsuan Hung, I-Fan Lin +3
In this study, we propose an encoder-decoder structured system with fully convolutional networks to implement voice activity detection (VAD) directly on the time-domain waveform. T…
Boosting Objective Scores of a Speech Enhancement Model by MetricGAN Post-processing
Szu-Wei Fu, Chien-Feng Liao, Tsun-An Hsieh +9
The Transformer architecture has demonstrated a superior ability compared to recurrent neural networks in many different natural language processing applications. Therefore, our st…
iMetricGAN: Intelligibility Enhancement for Speech-in-Noise using Generative Adversarial Network-based Metric Learning
Haoyu Li, Szu-Wei Fu, Yu Tsao +1
The intelligibility of natural speech is seriously degraded when exposed to adverse noisy environments. In this work, we propose a deep learning-based speech modification method to…