3 papers
cs.LG2024
Asymptotics of Language Model Alignment
Joy Qiping Yang, Salman Salamatian, Ziteng Sun +2
Let denote a generative language model. Let denote a reward model that returns a scalar that captures the degree at which a draw from is preferred. The goal of language…
cs.DS2023
Simpler Distribution Testing with Little Memory
Clément L. Canonne, Joy Qiping Yang
We consider the question of distribution testing (specifically, uniformity and closeness testing) in the streaming setting, \ie under stringent memory constraints. We improve on th…
cs.LG2023
Near-Optimal Degree Testing for Bayes Nets
Vipul Arora, Arnab Bhattacharyya, Clément L. Canonne +1
This paper considers the problem of testing the maximum in-degree of the Bayes net underlying an unknown probability distribution over , given sample access to .…