12 citations · 20 across the 6 of their papers we have counts for
4 papers
Beyond the Best: Estimating Distribution Functionals in Infinite-Armed Bandits
Yifei Wang, Tavor Baharav, Yanjun Han +2
In the infinite-armed bandit problem, each arm's average reward is sampled from an unknown distribution, and each arm can be sampled further to obtain noisy estimates of the averag…
Optimal Conservative Offline RL with General Function Approximation via Augmented Lagrangian
Paria Rashidinejad, Hanlin Zhu, Kunhe Yang +2
Offline reinforcement learning (RL), which refers to decision-making from a previously-collected dataset of interactions, has received significant attention over the past years. Mu…
Beyond Maximum Likelihood: from Theory to Practice
Jiantao Jiao, Kartik Venkat, Yanjun Han +1
Maximum likelihood is the most widely used statistical estimation technique. Recent work by the authors introduced a general methodology for the construction of estimators for func…
An Extremal Inequality for Long Markov Chains
Thomas Courtade, Jiantao Jiao
Let be jointly Gaussian vectors, and consider random variables that satisfy the Markov constraint . We prove an extremal inequality relating the mutual informa…