Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
FastMix: Fast Data Mixture Optimization via Gradient Descent
Haoru Tan, Sitong Wu, Yanfeng Chen +5
While large and diverse datasets have driven recent advances in large models, identifying the optimal data mixture for pre-training and post-training remains a significant open pro…
cs.LG2025
FairLRF: Achieving Fairness through Sparse Low Rank Factorization
Yuanbo Guo, Jun Xia, Yiyu Shi
As deep learning (DL) techniques become integral to various applications, ensuring model fairness while maintaining high performance has become increasingly critical, particularly…
cs.LG2025
NeFT: Negative Feedback Training to Improve Robustness of Compute-In-Memory DNN Accelerators
Yifan Qin, Zheyu Yan, Dailin Gan +5
Compute-in-memory accelerators built upon non-volatile memory devices excel in energy efficiency and latency when performing deep neural network (DNN) inference, thanks to their in…