11 citations · 15 across the 4 of their papers we have counts for
8 papers
Data-Driven Adaptive Simultaneous Machine Translation
Guangxu Xun, Mingbo Ma, Yuchen Bian +7
In simultaneous translation (SimulMT), the most widely used strategy is the wait-k policy thanks to its simplicity and effectiveness in balancing translation quality and latency. H…
Exploiting a Zoo of Checkpoints for Unseen Tasks
Jiaji Huang, Qiang Qiu, Kenneth Church
There are so many models in the literature that it is difficult for practitioners to decide which combinations are likely to be effective for a new task. This paper attempts to add…
Better than BERT but Worse than Baseline
Boxiang Liu, Jiaji Huang, Xingyu Cai +1
This paper compares BERT-SQuAD and Ab3P on the Abbreviation Definition Identification (ADI) task. ADI inputs a text and outputs short forms (abbreviations/acronyms) and long forms…
DiffWave: A Versatile Diffusion Model for Audio Synthesis
Zhifeng Kong, Wei Ping, Jiaji Huang +2
In this work, we propose DiffWave, a versatile diffusion probabilistic model for conditional and unconditional waveform generation. The model is non-autoregressive, and converts th…
Language Modeling at Scale
Mostofa Patwary, Milind Chabbi, Heewoo Jun +3
We show how Zipf's Law can be used to scale up language modeling (LM) to take advantage of more training data and more GPUs. LM plays a key role in many important natural language…
Large Margin Neural Language Model
Jiaji Huang, Yi Li, Wei Ping +1
We propose a large margin criterion for training neural language models. Conventionally, neural language models are trained by minimizing perplexity (PPL) on grammatical sentences.…