1 paper
Jiming Su, Hantao Hua, Lujia Yin +2
In simulation-in-the-loop decision-making systems, reinforcement learning (RL) inference is often constrained by simulator-side execution overhead, where workloads are highly dynam…