activity
20182020
most citedContext, Attention and Audio Feature Explorations for Audio Visual Scene-Aware Dialog

9 citations · 12 across the 6 of their papers we have counts for

collaborators

8 papers

eess.AS2020

Utterance-level Intent Recognition from Keywords

Wenda Chen, Jonathan Huang, Mark Hasegawa-Johnson

This paper focuses on wake on intent (WOI) techniques for platforms with limited compute and memory. Our approach of utterance-level intent classification is based on a sequence of…

eess.AS2020

Length- and Noise-aware Training Techniques for Short-utterance Speaker Recognition

Wenda Chen, Jonathan Huang, Tobias Bocklet

Speaker recognition performance has been greatly improved with the emergence of deep learning. Deep neural networks show the capacity to effectively deal with impacts of noise and…

eess.AS2020

Compact Speaker Embedding: lrx-vector

Munir Georges, Jonathan Huang, Tobias Bocklet

Deep neural networks (DNN) have recently been widely used in speaker recognition systems, achieving state-of-the-art performance on various benchmarks. The x-vector architecture is…

cs.CL20191 cited

Exploring Context, Attention and Audio Features for Audio Visual Scene-Aware Dialog

Shachi H Kumar, Eda Okur, Saurav Sahay +2

We are witnessing a confluence of vision, speech and dialog system technologies that are enabling the IVAs to learn audio-visual groundings of utterances and have conversations wit…

cs.MM20192 cited

Leveraging Topics and Audio Features with Multimodal Attention for Audio Visual Scene-Aware Dialog

Shachi H Kumar, Eda Okur, Saurav Sahay +2

With the recent advancements in Artificial Intelligence (AI), Intelligent Virtual Assistants (IVA) such as Alexa, Google Home, etc., have become a ubiquitous part of many homes. Cu…

cs.CL20189 cited

Context, Attention and Audio Feature Explorations for Audio Visual Scene-Aware Dialog

Shachi H Kumar, Eda Okur, Saurav Sahay +3

With the recent advancements in AI, Intelligent Virtual Assistants (IVA) have become a ubiquitous part of every home. Going forward, we are witnessing a confluence of vision, speec…