4 citations · 7 across the 3 of their papers we have counts for
4 papers
CCATMos: Convolutional Context-aware Transformer Network for Non-intrusive Speech Quality Assessment
Yuchen Liu, Li-Chia Yang, Alex Pawlicki +1
Speech quality assessment has been a critical component in many voice communication related applications such as telephony and online conferencing. Traditional intrusive speech qua…
Self-Supervised Learning for Speech Enhancement through Synthesis
Bryce Irvin, Marko Stamenovic, Mikolaj Kegler +1
Modern speech enhancement (SE) networks typically implement noise suppression through time-frequency masking, latent representation masking, or discriminative signal prediction. In…
Remixing Music with Visual Conditioning
Li-Chia Yang, Alexander Lerch
We propose a visually conditioned music remixing system by incorporating deep visual and audio models. The method is based on a state of the art audio-visual source separation mode…
Neural Wavetable: a playable wavetable synthesizer using neural networks
Lamtharn Hantrakul, Li-Chia Yang
We present Neural Wavetable, a proof-of-concept wavetable synthesizer that uses neural networks to generate playable wavetables. The system can produce new, distinct waveforms thro…