9 citations · 12 across the 6 of their papers we have counts for
8 papers
Utterance-level Intent Recognition from Keywords
Wenda Chen, Jonathan Huang, Mark Hasegawa-Johnson
This paper focuses on wake on intent (WOI) techniques for platforms with limited compute and memory. Our approach of utterance-level intent classification is based on a sequence of…
Length- and Noise-aware Training Techniques for Short-utterance Speaker Recognition
Wenda Chen, Jonathan Huang, Tobias Bocklet
Speaker recognition performance has been greatly improved with the emergence of deep learning. Deep neural networks show the capacity to effectively deal with impacts of noise and…
Compact Speaker Embedding: lrx-vector
Munir Georges, Jonathan Huang, Tobias Bocklet
Deep neural networks (DNN) have recently been widely used in speaker recognition systems, achieving state-of-the-art performance on various benchmarks. The x-vector architecture is…
Exploring Context, Attention and Audio Features for Audio Visual Scene-Aware Dialog
Shachi H Kumar, Eda Okur, Saurav Sahay +2
We are witnessing a confluence of vision, speech and dialog system technologies that are enabling the IVAs to learn audio-visual groundings of utterances and have conversations wit…
Leveraging Topics and Audio Features with Multimodal Attention for Audio Visual Scene-Aware Dialog
Shachi H Kumar, Eda Okur, Saurav Sahay +2
With the recent advancements in Artificial Intelligence (AI), Intelligent Virtual Assistants (IVA) such as Alexa, Google Home, etc., have become a ubiquitous part of many homes. Cu…
Context, Attention and Audio Feature Explorations for Audio Visual Scene-Aware Dialog
Shachi H Kumar, Eda Okur, Saurav Sahay +3
With the recent advancements in AI, Intelligent Virtual Assistants (IVA) have become a ubiquitous part of every home. Going forward, we are witnessing a confluence of vision, speec…