4 papers
Estimating Implicit Regularization in Deep Learning
Joseph H. Rudoler, Kevin Tan, Giles Hooker +1
Deep learning systems are known to exhibit implicit regularization (alt. implicit bias), favoring simple solutions instead of merely minimizing the loss function. In some cases, we…
Optimistic Actor-Critic with Parametric Policies for Linear Markov Decision Processes
Max Qiushi Lin, Reza Asad, Kevin Tan +3
Although actor-critic methods have been successful in practice, their theoretical analyses have several limitations. Specifically, existing theoretical work either sidesteps the ex…
Statistical Inference under Adaptive Sampling with LinUCB
Wei Fan, Kevin Tan, Yuting Wei
Adaptively collected data has become ubiquitous within modern practice. However, even seemingly benign adaptive sampling schemes can introduce severe biases, rendering traditional…
Actor-Critics Can Achieve Optimal Sample Efficiency
Kevin Tan, Wei Fan, Yuting Wei
Actor-critic algorithms have become a cornerstone in reinforcement learning (RL), leveraging the strengths of both policy-based and value-based methods. Despite recent progress in…