From the 1 of 1 linked paper with an AI index.
1 paper
Derek Ming Siang Tan, Shailesh Shailesh, Srikrishna Iyer +4
The paper presents RL², an adaptive test‑time steering framework that uses offline reinforcement learning on latent features from a frozen Vision‑Language‑Action model to compose a…