3.4k citations
- Google (United States)US56 papers
- Google (United Kingdom)GB17 papers
- University of TorontoCA7 papers
- Centre de Recherche en InformatiqueFR6 papers
- Centre de Recherche en Informatique, Signal et Automatique de LilleFR6 papers
- University of AlbertaCA6 papers
- University of OxfordGB6 papers
- Carnegie Mellon UniversityUS5 papers
- Columbia UniversityUS5 papers
- McGill UniversityCA5 papers
- Afterschool AllianceUS4 papers
- École Normale Supérieure - PSLFR4 papers
5 papers · 1 filter
Multi-view Subword Regularization
Xinyi Wang, Sebastian Ruder, Graham Neubig
Multilingual pretrained representations generally rely on subword segmentation algorithms to create a shared multilingual vocabulary. However, standard heuristic algorithms often l…
Decoupling the Role of Data, Attention, and Losses in Multimodal Transformers
Lisa Anne Hendricks, John Mellor, Rosalia Schneider +2
Recently multimodal transformer models have gained popularity because their performance on language and vision tasks suggest they learn rich visual-linguistic representations. Focu…
Improving Adversarial Text Generation by Modeling the Distant Future
Ruiyi Zhang, Changyou Chen, Zhe Gan +5
Auto-regressive text generation models usually focus on local fluency, and may cause inconsistent semantic meaning in long text generation. Further, automatically generating words…
Nested-Wasserstein Self-Imitation Learning for Sequence Generation
Ruiyi Zhang, Changyou Chen, Zhe Gan +3
Reinforcement learning (RL) has been widely studied for improving sequence-generation models. However, the conventional rewards used for RL training typically cannot capture suffic…
Listen and Translate: A Proof of Concept for End-to-End Speech-to-Text Translation
Alexandre Berard, Olivier Pietquin, Christophe Servan +1
This paper proposes a first attempt to build an end-to-end speech-to-text translation system, which does not use source language transcription during learning or decoding. We propo…