1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Hanchen Zhang, Xiao Liu, Bowen Lv +11
Recent advances in large language models (LLMs) have sparked growing interest in building generalist agents that can learn through online interactions. However, applying reinforcem…