Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Prior Diffusiveness and Regret in the Linear-Gaussian Bandit
Yifan Zhu, John C. Duchi, Benjamin Van Roy
We prove that Thompson sampling exhibits Bayesian regret in the linear-Gaussian bandit with a pri…
cs.LG2026
A Bitter Lesson for Data Filtering
Christopher Mohri, John Duchi, Tatsunori Hashimoto
We investigate data filtering for large model pretraining via new scaling studies that target the high compute, data-scarce regime. In spite of an apparently common belief that fil…