Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
On the Convergence Analysis of Muon
Wei Shen, Ruichuan Huang, Minhui Huang +2
The majority of parameters in neural networks are naturally represented as matrices. However, most commonly used optimizers treat these matrix parameters as flattened vectors durin…
stat.ML2024
Stochastic Smoothed Gradient Descent Ascent for Federated Minimax Optimization
Wei Shen, Minhui Huang, Jiawei Zhang +1
In recent years, federated minimax optimization has attracted growing interest due to its extensive applications in various machine learning tasks. While Smoothed Alternative Gradi…