2 papers
cs.RO2026
SafeStage: Evaluating Safety Before, During, and After Vision-Language-Conditioned Robot Manipulation
Jinzhu Luo, Qi Zhang, Wei Wang +1
Vision-language-conditioned robot policies integrate perception, language understanding, and control for general-purpose manipulation. However, existing evaluations often focus on…
cs.LG2025
An Empirical Study on the Effectiveness of Incorporating Offline RL As Online RL Subroutines
Jianhai Su, Jinzhu Luo, Qi Zhang
We take the novel perspective of incorporating offline RL algorithms as subroutines of tabula rasa online RL. This is feasible because an online learning agent can repurpose its hi…