Training Neural Response Selection for Task-Oriented Dialogue Systems
arXiv:1906.01543
Abstract
Despite their popularity in the chatbot literature, retrieval-based models have had modest impact on task-oriented dialogue systems, with the main obstacle to their application being the low-data regime of most task-oriented dialogue tasks. Inspired by the recent success of pretraining in language modelling, we propose an effective method for deploying response selection in task-oriented dialogue. To train response selection models for task-oriented dialogue tasks, we propose a novel method which: 1) pretrains the response selection model on large general-domain conversational corpora; and then 2) fine-tunes the pretrained model for the target dialogue domain, relying only on the small in-domain dataset to capture the nuances of the given dialogue domain. Our evaluation on six diverse application domains, ranging from e-commerce to banking, demonstrates the effectiveness of the proposed training method.
ACL 2019 long paper
References in corpus (5)
- Self-Normalizing Neural Networks
- Regularizing Neural Networks by Penalizing Confident Output Distributions
- An Information Retrieval Approach to Short Text Conversation
- Simple Applications of BERT for Ad Hoc Document Retrieval
- Scale out for large minibatch SGD: Residual network training on ImageNet-1K with improved accuracy and reduced time to train