2 citations · 2 across the 1 of their papers we have counts for
1 paper
Haozhe Ji, Pei Ke, Zhipeng Hu +2
The standard paradigm of neural language generation adopts maximum likelihood estimation (MLE) as the optimizing method. From a distributional view, MLE in fact minimizes the Kullb…