Showing stat.MLShow all
2 papers · 1 filter
stat.ML2025
Best-of-N through the Smoothing Lens: KL Divergence and Regret Analysis
Gholamali Aminian, Idan Shenfeld, Amir R. Asadi +2
A simple yet effective method for inference-time alignment of generative models is Best-of- (BoN), where outcomes are sampled from a reference policy, evaluated using a prox…
stat.ML2025
Generalization and Robustness of the Tilted Empirical Risk
Gholamali Aminian, Amir R. Asadi, Tian Li +3
The generalization error (risk) of a supervised statistical learning algorithm quantifies its prediction ability on previously unseen data. Inspired by exponential tilting, \citet{…