diffusion models 1human-in-the-loop 1policy adaptation 1reinforcement learning 1vision-language-action 1
From the 1 of 5 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Reinforcing VLAs in Task-Agnostic World Models
Yucen Wang, Rui Yu, Fengming Zhang +5
Post-training Vision-Language-Action (VLA) models via reinforcement learning (RL) in learned world models has emerged as an effective strategy to adapt to new tasks without costly…
cs.AI2025
AdaptiveStep: Automatically Dividing Reasoning Step through Model Confidence
Yuliang Liu, Junjie Lu, Zhaoling Chen +10
Current approaches for training Process Reward Models (PRMs) often involve breaking down responses into multiple reasoning steps using rule-based techniques, such as using predefin…