2 papers
math.ST2026
Estimation of the sub-Gaussian parameter
Jason Liu, Min Xu, Jinchuan Xing
The sub-Gaussian parameter (also called the variance proxy) of a mean-zero random variable is defined as where $L(λ) = \frac{2}{λ^2}…
cs.LG2026
Enhancing the MADDPG Algorithm for Multi-Agent Learning via Action Inference and Importance Sampling
Marc Walden, Jason Liu, Shaashwath Sivakumar +2
We investigate multi-agent deep reinforcement learning and propose two enhancements to the Multi-Agent Deep Deterministic Policy Gradient (MADDPG) algorithm. First, we introduce a…