Neural Machine Translation and Sequence-to-sequence Models: A Tutorial
arXiv:1703.01619
Abstract
This tutorial introduces a new and powerful set of techniques variously called "neural machine translation" or "neural sequence-to-sequence models". These techniques have been used in a number of tasks regarding the handling of human language, and can be a powerful tool in the toolbox of anyone who wants to model sequential data of some sort. The tutorial assumes that the reader knows the basics of math and programming, but does not assume any particular experience with neural networks or natural language processing. It attempts to explain the intuition behind the various methods covered, then delves into them with enough mathematical detail to understand them concretely, and culiminates with a suggestion for an implementation exercise, where readers can test that they understood the content in practice.
65 Pages
References in corpus (11)
- Sequence to Sequence Learning with Neural Networks
- Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling
- Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
- Neural Architecture Search with Reinforcement Learning
- From Frequency to Meaning: Vector Space Models of Semantics
- DRAW: A Recurrent Neural Network For Image Generation
- Transition-Based Dependency Parsing with Stack Long Short-Term Memory
- DyNet: The Dynamic Neural Network Toolkit
- Neural Machine Translation in Linear Time
- Toward Multilingual Neural Machine Translation with Universal Encoder and Decoder
- Reward Augmented Maximum Likelihood for Neural Structured Prediction