2 papers
cs.MA2026
Generalized Per-Agent Advantage Estimation for Multi-Agent Policy Optimization
Seongmin Kim, Giseung Park, Woojun Kim +3
In this paper, we propose a novel framework for multi-agent reinforcement learning that enhances sample efficiency and coordination through accurate per-agent advantage estimation.…
math.OC2025
Fixed Confidence and Fixed Tolerance Bi-level Optimization for Selecting the Best Optimized System
Yuhao Wang, Seong-Hee Kim, Enlu Zhou
In this paper, we study a fixed-confidence, fixed-tolerance formulation of a class of stochastic bi-level optimization problems, where the upper-level problem selects from a finite…