Showing stat.MLShow all
2 papers · 1 filter
stat.ML2025
Learning Guarantee of Reward Modeling Using Deep Neural Networks
Yuanhang Luo, Yeheng Ge, Ruijian Han +1
In this work, we study the learning theory of reward modeling with pairwise comparison data using deep neural networks. We establish a novel non-asymptotic regret bound for deep re…
stat.ML2025
Adaptive debiased SGD in high-dimensional GLMs with streaming data
Ruijian Han, Lan Luo, Yuanhang Luo +2
Online statistical inference facilitates real-time analysis of sequentially collected data, making it different from traditional methods that rely on static datasets. This paper in…