2 papers
cs.LG2026
OmniOpt: Taxonomy, Geometry, and Benchmarking of Modern Optimizers
Siyuan Li, Jiabao Pan, Yumou Liu +9
Optimizer selection for large-scale model training has become a system-level design decision constrained jointly by compute, memory, tuning budget, and task diversity, yet the land…
cs.LG2026
De-attribute to Forget for LLM Unlearning
Xinyang Lu, Jiabao Pan, Rachael Hwee Ling Sim +3
The rapid development of large language models (LLMs) has raised concerns on the use of inappropriate data for training, which has led to a growing interest in LLM unlearning. Many…