1 paper
Yijing Ke, Zihan Zhang, Ruosong Wang
We study computationally and statistically efficient reinforcement learning under the linear QI¨ realizability assumption, where any policy's Q-function is linear in a given s…