3 papers
cs.LG2025
A Modular Algorithm for Non-Stationary Online Convex-Concave Optimization
Qing-xin Meng, Xia Lei, Jian-wei Liu
This paper investigates the problem of Online Convex-Concave Optimization, which extends Online Convex Optimization to two-player time-varying convex-concave games. The goal is to…
cs.CV2025
Reinforcement Learning for Large Model: A Survey
Weijia Wu, Chen Gao, Joya Chen +6
Recent advances at the intersection of reinforcement learning (RL) and visual intelligence have enabled agents that not only perceive complex visual scenes but also reason, generat…
cs.LG2024
Proximal Point Method for Online Saddle Point Problem
Qing-xin Meng, Jian-wei Liu
This paper focuses on the online saddle point problem, which involves a sequence of two-player time-varying convex-concave games. Considering the nonstationarity of the environment…