2 papers
cs.LG2024
Learning from Sparse Offline Datasets via Conservative Density Estimation
Zhepeng Cen, Zuxin Liu, Zitong Wang +3
Offline reinforcement learning (RL) offers a promising direction for learning policies from pre-collected datasets without requiring further interactions with the environment. Howe…
stat.ML2023
Cheap Bootstrap for Fast Uncertainty Quantification of Stochastic Gradient Descent
Henry Lam, Zitong Wang
Stochastic gradient descent (SGD) or stochastic approximation has been widely used in model training and stochastic optimization. While there is a huge literature on analyzing its…