3 papers
eess.AS2021
Streaming Transformer for Hardware Efficient Voice Trigger Detection and False Trigger Mitigation
Vineet Garg, Wonil Chang, Siddharth Sigtia +4
We present a unified and hardware efficient architecture for two stage voice trigger detection (VTD) and false trigger mitigation (FTM) tasks. Two stage VTD systems of voice assist…
cs.LG2019
Orthogonality Constrained Multi-Head Attention For Keyword Spotting
Mingu Lee, Jinkyu Lee, Hye Jin Jang +3
Multi-head attention mechanism is capable of learning various representations from sequential data while paying attention to different subsequences, e.g., word-pieces or syllables…
eess.AS2019
An End-to-End Text-independent Speaker Verification Framework with a Keyword Adversarial Network
Sungrack Yun, Janghoon Cho, Jungyun Eum +2
This paper presents an end-to-end text-independent speaker verification framework by jointly considering the speaker embedding (SE) network and automatic speech recognition (ASR) n…