latent compositional steering 1offline reinforcement learning 1out-of-domain generalization 1test-time adaptation 1vision-language-action 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.RO2026
RL-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models
Derek Ming Siang Tan, Shailesh Shailesh, Srikrishna Iyer +4
The paper presents RL², an adaptive test‑time steering framework that uses offline reinforcement learning on latent features from a frozen Vision‑Language‑Action model to compose a…
cs.AI2025
Application and Evaluation of Large Language Models for Forecasting the Impact of Traffic Incidents
George Jagadeesh, Srikrishna Iyer, Michal Polanowski +1
This study examines the feasibility of applying large language models (LLMs) for forecasting the impact of traffic incidents on the traffic flow. The use of LLMs for this task has…
cs.CL2024
When Babies Teach Babies: Can student knowledge sharing outperform Teacher-Guided Distillation on small datasets?
Srikrishna Iyer
We present our submission to the BabyLM challenge, aiming to push the boundaries of data-efficient language model pretraining. Our method builds upon deep mutual learning, introduc…