26 citations · 34 across the 6 of their papers we have counts for
6 papers
A Pre-training Framework that Encodes Noise Information for Speech Quality Assessment
Subrina Sultana, Donald S. Williamson
Self-supervised learning (SSL) has grown in interest within the speech processing community, since it produces representations that are useful for many downstream tasks. SSL uses g…
The Impact of Perceived Tone, Age, and Gender on Voice Assistant Persuasiveness in the Context of Product Recommendations
Sabid Bin Habib Pias, Ran Huang, Donald Williamson +2
Voice Assistants (VAs) can assist users in various everyday tasks, but many users are reluctant to rely on VAs for intricate tasks like online shopping. This study aims to examine…
Privacy-preserving and Privacy-attacking Approaches for Speech and Audio -- A Survey
Yuchen Liu, Apu Kapadia, Donald Williamson
In contemporary society, voice-controlled devices, such as smartphones and home assistants, have become pervasive due to their advanced capabilities and functionality. The always-o…
MMViT: Multiscale Multiview Vision Transformers
Yuchen Liu, Natasha Ong, Kaiyan Peng +8
We present Multiscale Multiview Vision Transformers (MMViT), which introduces multiscale feature maps and multiview encodings to transformer models. Our model encodes different vie…
Attention-based Speech Enhancement Using Human Quality Perception Modelling
Khandokar Md. Nayem, Donald S. Williamson
Perceptually-inspired objective functions such as the perceptual evaluation of speech quality (PESQ), signal-to-distortion ratio (SDR), and short-time objective intelligibility (ST…
A Composite T60 Regression and Classification Approach for Speech Dereverberation
Yuying Li, Yuchen Liu, Donald S. Williamson
Dereverberation is often performed directly on the reverberant audio signal, without knowledge of the acoustic environment. Reverberation time, T60, however, is an essential acoust…