3 papers
cs.CL2026
Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models
Feng Luo, Yu-Neng Chuang, Guanchu Wang +4
On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify a failure mode of OPD: as t…
cs.LG2025
FairAgent: Democratizing Fairness-Aware Machine Learning with LLM-Powered Agents
Yucong Dai, Lu Zhang, Feng Luo +2
Training fair and unbiased machine learning models is crucial for high-stakes applications, yet it presents significant challenges. Effective bias mitigation requires deep expertis…
cs.LG2025
Integrating Fairness and Model Pruning Through Bi-level Optimization
Yucong Dai, Gen Li, Feng Luo +2
Deep neural networks have achieved exceptional results across a range of applications. As the demand for efficient and sparse deep learning models escalates, the significance of mo…