Showing eess.ASShow all
3 papers · 1 filter
eess.AS2025
Speaker-IPL: Unsupervised Learning of Speaker Characteristics with i-Vector based Pseudo-Labels
Zakaria Aldeneh, Takuya Higuchi, Jee-weon Jung +6
Iterative self-training, or iterative pseudo-labeling (IPL) -- using an improved model from the current iteration to provide pseudo-labels for the next iteration -- has proven to b…
eess.AS2024
Does Single-channel Speech Enhancement Improve Keyword Spotting Accuracy? A Case Study
Avamarie Brueggeman, Takuya Higuchi, Masood Delfarah +2
Noise robustness is a key aspect of successful speech applications. Speech enhancement (SE) has been investigated to improve automatic speech recognition accuracy; however, its eff…
eess.AS2024
Multichannel Voice Trigger Detection Based on Transform-average-concatenate
Takuya Higuchi, Avamarie Brueggeman, Masood Delfarah +1
Voice triggering (VT) enables users to activate their devices by just speaking a trigger phrase. A front-end system is typically used to perform speech enhancement and/or separatio…