Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge
Guangchen Lan, Sipeng Zhang, Tianle Wang +7
As the era of large language models (LLMs) unfolds, Preference Optimization (PO) methods have become a central approach to aligning LLMs with human preferences and improving perfor…
cs.LG2025
Gradient Correction in Federated Learning with Adaptive Optimization
Evan Chen, Shiqiang Wang, Jianing Zhang +3
In federated learning (FL), model training performance is strongly impacted by data heterogeneity across clients. Client-drift compensation methods have recently emerged as a solut…