3 papers
cs.LG2026
AdaMemento: Adaptive Memory-Assisted Policy Optimization for Reinforcement Learning
Renye Yan, Yaozhong Gan, You Wu +4
In sparse reward scenarios of reinforcement learning (RL), the memory mechanism provides promising shortcuts to policy optimization by reflecting on past experiences like humans. H…
cs.MA2025
MARPO: A Reflective Policy Optimization for Multi Agent Reinforcement Learning
Cuiling Wu, Yaozhong Gan, Junliang Xing +1
We propose Multi Agent Reflective Policy Optimization (MARPO) to alleviate the issue of sample inefficiency in multi agent reinforcement learning. MARPO consists of two key compone…
cs.CV2025
A Sanity Check for Multi-In-Domain Face Forgery Detection in the Real World
Jikang Cheng, Renye Yan, Zhiyuan Yan +5
Existing methods for deepfake detection aim to develop generalizable detectors. Although "generalizable" is the ultimate target once and for all, with limited training forgeries an…