1 citations · 1 across the 4 of their papers we have counts for
1 paper · 1 filter
Jialian Li, Tongzheng Ren, Dong Yan +2
In high-stake scenarios like medical treatment and auto-piloting, it's risky or even infeasible to collect online experimental data to train the agent. Simulation-based training ca…