11 citations · 34 across the 21 of their papers we have counts for
Showing 2026Show all
2 papers · 1 filter
cs.LG2026
Prompt-Driven Exploration: Language as an Exploration Space for VLA Reinforcement Learning
Sunshine Jiang, John Marangola, David Zhang +6
Exploration is essential to RL since a policy cannot improve by repeatedly sampling the behaviors it already prefers. Standard methods inject stochasticity in the action space, but…
cs.LG2026
Learning More from Less: Reinforcement Learning from Hindsight
Iris Xu, Sunshine Jiang, John Marangola +8
Reinforcement learning (RL) is increasingly used to post-train vision-language-action (VLA) models, but every update consumes robot rollouts that are slow and costly to collect, ma…