1 paper · 1 filter
Wonje Jeung, Sangyeon Yoon, Hyesoo Hong +6
Vision-language models are increasingly used as reward functions for robotic learning, but this role requires paraphrase invariance: the same trajectory should receive the same rew…