9 citations · 9 across the 1 of their papers we have counts for
2 papers
cs.CL2018★ 9 cited
Context, Attention and Audio Feature Explorations for Audio Visual Scene-Aware Dialog
Shachi H Kumar, Eda Okur, Saurav Sahay +3
With the recent advancements in AI, Intelligent Virtual Assistants (IVA) have become a ubiquitous part of every home. Going forward, we are witnessing a confluence of vision, speec…
cs.SD2018
AclNet: efficient end-to-end audio classification CNN
Jonathan J Huang, Juan Jose Alvarado Leanos
We propose an efficient end-to-end convolutional neural network architecture, AclNet, for audio classification. When trained with our data augmentation and regularization, we achie…