2 papers
cs.LG2025
AdamHD: Decoupled Huber Decay Regularization for Language Model Pre-Training
Fu-Ming Guo, Yingfang Fan
Adaptive optimizers with decoupled weight decay, such as AdamW, are the de facto standard for pre-training large transformer-based generative models. Yet the quadratic nature of th…
cs.IR2025
StealthRank: LLM Ranking Manipulation via Stealthy Prompt Optimization
Yiming Tang, Yi Fan, Chenxiao Yu +3
The integration of large language models (LLMs) into information retrieval systems introduces new attack surfaces, particularly for adversarial ranking manipulations. We present $\…