Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Feedback Manipulation Regularization: Enabling Offline Agent Alignment for Imitation Learning
Benjamin Poole, Minwoo Lee
Reinforcement learning (RL) research has increasingly shifted focus towards alignment, ensuring agents learn behaviors adhering to human values. While human demonstrations and feed…
cs.AI2026
From Static Risk to Dynamic Trajectories: Toward World-Model-Inspired Clinical Prediction
Pujun Feng, Xiaoyu Guo, Seyed Ehsan Saffari +10
Clinical decision-making is a feedback system where risk estimates influence treatment, which in turn changes disease trajectories, and both shape clinicians' measurement practices…