Hierarchical Text Generation and Planning for Strategic Dialogue
arXiv:1712.05846
Abstract
End-to-end models for goal-orientated dialogue are challenging to train, because linguistic and strategic aspects are entangled in latent state vectors. We introduce an approach to learning representations of messages in dialogues by maximizing the likelihood of subsequent sentences and actions, which decouples the semantics of the dialogue utterance from its linguistic realization. We then use these latent sentence representations for hierarchical language generation, planning and reinforcement learning. Experiments show that our approach increases the end-task reward achieved by the model, improves the effectiveness of long-term planning using rollouts, and allows self-play reinforcement learning to improve decision making without diverging from human language. Our hierarchical latent-variable model outperforms previous work both linguistically and strategically.
Cited by in corpus (16)
- Hybrid Retrieval-Generation Reinforced Agent for Medical Image Report Generation
- Hierarchical Decision Making by Generating and Following Natural Language Instructions
- Modelling Hierarchical Structure between Dialogue Policy and Natural Language Generator with Option Framework for Task-oriented Dialogue System
- MALA: Cross-Domain Dialogue Generation with Action Learning
- Compositional Transformers for Scene Generation
- Recommendation as a Communication Game: Self-Supervised Bot-Play for Goal-oriented Dialogue
- Exclusive Hierarchical Decoding for Deep Keyphrase Generation
- Judge the Judges: A Large-Scale Evaluation Study of Neural Language Models for Online Review Generation
- Document-editing Assistants and Model-based Reinforcement Learning as a Path to Conversational AI
- Know More about Each Other: Evolving Dialogue Strategy via Compound Assessment
- Learning Goal-oriented Dialogue Policy with Opposite Agent Awareness
- Dynamic Knowledge Routing Network For Target-Guided Open-Domain Conversation
- Generalizable and Explainable Dialogue Generation via Explicit Action Learning
- Maintaining Common Ground in Dynamic Environments
- Targeted Data Acquisition for Evolving Negotiation Agents
- Target-Guided Open-Domain Conversation