1 paper
Mahyar Alinejad, Alvaro Velasquez, Yue Wang +1
Reinforcement Learning (RL) in environments with complex, history-dependent reward structures poses significant challenges for traditional methods. In this work, we introduce a nov…