5 citations · 5 across the 3 of their papers we have counts for
1 paper · 1 filter
Feihu Jin, Ying Tan
Fine-tuning large language models (LLMs) using standard first-order (FO) optimization often drives training toward sharp, poorly generalizing minima. Conversely, zeroth-order (ZO)…