2 papers
cs.AI2026
RetroAgent: From Solving to Evolving via Retrospective Dual Intrinsic Feedback
Xiaoying Zhang, Zichen Liu, Yipeng Zhang +2
Standard reinforcement learning (RL) for large language model (LLM) agents primarily optimizes extrinsic task rewards, often favoring isolated task completion over continual adapta…
math.OC2026
A Note on the Gradient-Evaluation Sequence in Accelerated Gradient Methods
Yan Wu, Yipeng Zhang, Lu Liu +1
Nesterov's accelerated gradient descent method (AGD) is a seminal deterministic first-order method known to achieve the optimal order of iteration complexity for solving convex smo…