9 citations · 19 across the 7 of their papers we have counts for
1 paper · 1 filter
Alexander Lin, Jeremy Wohlwend, Howard Chen +1
The performance of autoregressive models on natural language generation tasks has dramatically improved due to the adoption of deep, self-attentive architectures. However, these ga…