Fast Rates in -Potential Games via Regularized Mirror Descent
arXiv:2605.00268
Abstract
An -potential game is a multi-player non-cooperative interaction in which a global potential function approximates individual player rewards up to a structural bias . While identifying a Nash Equilibrium (NE) in generic general-sum games is known to be computationally intractable, the potential game structure enables tractable NE identification. In this paper, we study the offline learning of NE in -potential games using KL regularization. To analyze this process, we propose a novel Reference-Anchored offline data coverage framework--a verifiable condition that anchors data requirements to a known reference policy rather than an unknown optimum. Building on this, we propose Offline Potential Mirror Descent (OPMD), a decentralized algorithm that achieves an accelerated statistical rate, surpassing the standard rate typical of offline multi-agent learning. This work characterizes the first fast-rate offline learning approach for -potential games.