1 paper · 1 filter
He Wang, Laixi Shi, Yuejie Chi
In offline reinforcement learning (RL), the absence of active exploration calls for attention on the model robustness to tackle the sim-to-real gap, where the discrepancy between t…