1 paper
Shurong Mo, Nailong Wu, Jie Qi +4
This study proposes a delay-compensated feedback controller based on proximal policy optimization (PPO) reinforcement learning to stabilize traffic flow in the congested regime by…