1 paper
Zen Kit Heng, Zimeng Zhao, Tianhao Wu +4
Large Language Models (LLMs) are emerging as promising tools for automated reinforcement learning (RL) reward design, owing to their robust capabilities in commonsense reasoning an…