3 citations · 11 across the 16 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
BRAIn: Bayesian Reward-conditioned Amortized Inference for natural language generation from feedback
Gaurav Pandey, Yatin Nandwani, Tahira Naseem +6
Distribution matching methods for language model alignment such as Generation with Distributional Control (GDC) and Distributional Policy Gradient (DPG) have not received the same…
cs.LG2016
An improved uncertainty decoding scheme with weighted samples for DNN-HMM hybrid systems
Christian Huemmer, Ramón Fernández Astudillo, Walter Kellermann
In this paper, we advance a recently-proposed uncertainty decoding scheme for DNN-HMM (deep neural network - hidden Markov model) hybrid systems. This numerical sampling concept av…