1 paper · 1 filter
Daiqi Gao, Ziping Xu, Aseel Rawashdeh +2
Measuring states in reinforcement learning (RL) can be costly in real-world settings and may negatively influence future outcomes. We introduce the Actively Observable Markov Decis…