1 citations · 1 across the 2 of their papers we have counts for
5 papers
Correct Online Estimation of the Powertrain Time Constants in Adaptive Vehicular Platooning
Qiuhao Wen, Simone Baldi, Jiwei Wang +2
In longitudinal platooning, some key sources of uncertainty are the powertrain time constants of the vehicles. Because such time constants appear in the input matrix of the platoon…
Enforcing Opacity in Discrete Event Systems via Delayed Observations
Jiwei Wang, Simone Baldi, Wenwu Yu +1
Artificially introducing a delay in the observations of a system can be an effective mechanism to mask the system itself, with the goal to increase its opacity and thus its securit…
Quantile Q-Learning: Revisiting Offline Extreme Q-Learning with Quantile Regression
Xinming Gao, Shangzhe Li, Yujin Cai +1
Offline reinforcement learning (RL) enables policy learning from fixed datasets without further environment interaction, making it particularly valuable in high-risk or costly doma…
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies
Pengcheng Dai, He Wang, Dongming Wang +1
This paper investigates constrained multi-agent reinforcement learning (CMARL) in coupled environments, where agents collaboratively maximize the sum of local objectives while sati…
Distributed Neural Policy Gradient Algorithm for Global Convergence of Networked Multi-Agent Reinforcement Learning
Pengcheng Dai, Yuanqiu Mo, Wenwu Yu +1
This paper studies the networked multi-agent reinforcement learning (NMARL) problem, where the objective of agents is to collaboratively maximize the discounted average cumulative…