3 papers
cs.CV2025
Reinforcement Learning for Large Model: A Survey
Weijia Wu, Chen Gao, Joya Chen +6
Recent advances at the intersection of reinforcement learning (RL) and visual intelligence have enabled agents that not only perceive complex visual scenes but also reason, generat…
cs.LG2025
A Modular Algorithm for Non-Stationary Online Convex-Concave Optimization
Qing-xin Meng, Xia Lei, Jian-wei Liu
This paper investigates the problem of Online Convex-Concave Optimization, which extends Online Convex Optimization to two-player time-varying convex-concave games. The goal is to…
cs.LG2025
Proximal Point Method for Online Saddle Point Problem
Qing-xin Meng, Jian-wei Liu
This paper focuses on the online saddle point problem, which involves a sequence of two-player time-varying convex-concave games. Considering the nonstationarity of the environment…