3 citations · 11 across the 14 of their papers we have counts for
5 papers · 1 filter
A Multimodal Target-Source Classifier with Attention Branches to Understand Ambiguous Instructions for Fetching Daily Objects
Aly Magassouba, Komei Sugiura, Hisashi Kawai
In this study, we focus on multimodal language understanding for fetching instructions in the domestic service robots context. This task consists of predicting a target object, as…
Cross-scale Attention Model for Acoustic Event Classification
Xugang Lu, Peng Shen, Sheng Li +2
A major advantage of a deep convolutional neural network (CNN) is that the focused receptive field size is increased by stacking multiple convolutional layers. Accordingly, the mod…
Multimodal Attention Branch Network for Perspective-Free Sentence Generation
Aly Magassouba, Komei Sugiura, Hisashi Kawai
In this paper, we address the automatic sentence generation of fetching instructions for domestic service robots. Typical fetching commands such as "bring me the yellow toy from th…
Understanding Natural Language Instructions for Fetching Daily Objects Using GAN-Based Multimodal Target-Source Classification
Aly Magassouba, Komei Sugiura, Anh Trinh Quoc +1
In this paper, we address multimodal language understanding for unconstrained fetching instruction in domestic service robots context. A typical fetching instruction such as "Bring…
Incorporating Symbolic Sequential Modeling for Speech Enhancement
Chien-Feng Liao, Yu Tsao, Xugang Lu +1
In a noisy environment, a lossy speech signal can be automatically restored by a listener if he/she knows the language well. That is, with the built-in knowledge of a "language mod…