2 papers
cs.LG2026
Flow Matching Policy Optimization with Mirror Descent and Entropy Constraints
Ting Gao, Stavros Orfanoudakis, Nan Lin +3
Balancing policy expressiveness with the exploration-exploitation trade-off is a core challenge in online Reinforcement Learning (RL). While Stochastic Differential Equation (SDE)-…
cs.LG2025
Q-Net: Queue Length Estimation via Kalman-based Neural Networks
Ting Gao, Elvin Isufi, Winnie Daamen +2
Estimating queue lengths at signalized intersections is a long-standing challenge in traffic management. Partial observability of vehicle flows complicates this task despite the av…