1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Yansong Qu, Zilin Huang, Zihao Sheng +4
Autonomous driving policy learning with reinforcement learning (RL) is fundamentally limited by low sample efficiency, weak generalization, and a dependence on unsafe online trial-…