large language model training 1matrix representations 1momentum methods 1optimizer geometry 1stochastic convergence 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.SE2026
PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization
Ryan Deng, Yuanzhe Liu, Bastian Lipka +4
Large language model (LLM) agents now perform well on correctness-oriented repository-level tasks, including SWE-Bench issue resolution and feature implementation in real codebases…
cs.LG2026
Muse: Representation Geometry of Muon Beyond Normalized Momentum
Da Chang, Qiankun Shi, Lvgang Zhang +4
The paper investigates how the choice of matrix representation influences Muon-style optimizers, proposes the Muse family of optimizers that keep the same momentum and Newton–Schul…