Showing math.OCShow all
2 papers · 1 filter
math.OC2026
Zeroth-Order Nonsmooth Nonconvex Optimization with Convex Liftings and Its Application to State-Feedback Policy Optimization
Xuhao Wang, Yujie Tang
Direct policy optimization is widely used in reinforcement learning and control, but generally leads to nonconvex optimization problems. For state-feedback control, the…
math.OC2026
Harnessing Individual Motivation for Collective Efficiency: A Mechanism-Driven Distributed Optimization Method
Dongwei Xie, Xuhao Wang, Yujie Tang +1
In industrial scenarios involving multi-agent collective decision-making, centralized decision-making may not be admissible due to restrictive access to individual local informatio…