1 paper
Akane Tsuboya, Yu Kono, Tatsuji Takahashi
The objective of a reinforcement learning agent is to discover better actions through exploration. However, typical exploration techniques aim to maximize rewards, often incurring…