1 paper
Yuxuan Lu, Ziyi Wang, Jing Huang +10
Reinforcement learning (RL) for web agents demands environments that are both effective for evaluation and efficient enough for large-scale on-policy training. Current web environm…