2 papers
cs.LG2026
Convergence of Multiagent Learning Systems for Traffic control
Sayambhu Sen, Shalabh Bhatnagar
Rapid urbanization in cities like Bangalore has led to severe traffic congestion, making efficient Traffic Signal Control (TSC) essential. Multi-Agent Reinforcement Learning (MARL)…
cs.LG2026
Enabling Off-Policy Imitation Learning with Deep Actor Critic Stabilization
Sayambhu Sen, Shalabh Bhatnagar
Learning complex policies with Reinforcement Learning (RL) is often hindered by instability and slow convergence, a problem exacerbated by the difficulty of reward engineering. Imi…