2 papers
cs.LG2026
Effective Quantization of Muon Optimizer States
Aman Gupta, Rafael Celente, Abhishek Shivanna +7
The Muon optimizer, based on matrix orthogonalization, has recently shown faster convergence and better computational efficiency over AdamW in LLM pre-training. However, the memory…
cs.IR2025
Your Spending Needs Attention: Modeling Financial Habits with Transformers
D. T. Braithwaite, Misael Cavalcanti, R. Austin McEver +9
Predictive models play a crucial role in the financial industry, enabling risk prediction, fraud detection, and personalized recommendations, where slight changes in core model per…