1 paper
Chengyu Luan, Bo Xin, Songyan Guo +4
Large language models can intervene in reinforcement learning through both reward design and action selection, yet aggregate performance offers an incomplete account of what these…