3 papers
cs.AI2026
Modeling Clinical Concern Trajectories in Language Model Agents
Sukesh Subaharan, Venkatesan VS, Murugadasan P +3
Large language model (LLM) agents deployed in clinical settings often exhibit abrupt, threshold-driven behavior, offering little visibility into accumulating risk prior to escalati…
cs.LG2026
Dynamical Priors as a Training Objective in Reinforcement Learning
Sukesh Subaharan
Standard reinforcement learning (RL) optimizes policies for reward but imposes few constraints on how decisions evolve over time. As a result, policies may achieve high performance…
cs.AI2026
Controlling Long-Horizon Behavior in Language Model Agents with Explicit State Dynamics
Sukesh Subaharan
Large language model (LLM) agents often exhibit abrupt shifts in tone and persona during extended interaction, reflecting the absence of explicit temporal structure governing agent…