25 citations · 51 across the 6 of their papers we have counts for
5 papers · 1 filter
Video-Grounded Dialogues with Pretrained Generation Language Models
Hung Le, Steven C. H. Hoi
Pre-trained language models have shown remarkable success in improving various downstream NLP tasks due to their ability to capture dependencies in textual data and generate natura…
UniConv: A Unified Conversational Neural Architecture for Multi-domain Task-oriented Dialogues
Hung Le, Doyen Sahoo, Chenghao Liu +2
Building an end-to-end conversational agent for multi-domain task-oriented dialogues has been an open challenge for two main reasons. First, tracking dialogue states of multiple do…
Multimodal Transformer with Pointer Network for the DSTC8 AVSD Challenge
Hung Le, Nancy F. Chen
Audio-Visual Scene-Aware Dialog (AVSD) is an extension from Video Question Answering (QA) whereby the dialogue agent is required to generate natural language responses to address u…
Non-Autoregressive Dialog State Tracking
Hung Le, Richard Socher, Steven C. H. Hoi
Recent efforts in Dialogue State Tracking (DST) for task-oriented dialogues have progressed toward open-vocabulary or generation-based approaches where the models can generate slot…
Multimodal Transformer Networks for End-to-End Video-Grounded Dialogue Systems
Hung Le, Doyen Sahoo, Nancy F. Chen +1
Developing Video-Grounded Dialogue Systems (VGDS), where a dialogue is conducted based on visual and audio aspects of a given video, is significantly more challenging than traditio…