3 papers
cs.CR2026
FedDetox: Robust Federated SLM Alignment via On-Device Data Sanitization
Shunan Zhu, Jiawei Chen, Yonghao Yu +1
As high quality public data becomes scarce, Federated Learning (FL) provides a vital pathway to leverage valuable private user data while preserving privacy. However, real-world cl…
cs.CE2025
Task-Specific Sparse Feature Masks for Molecular Toxicity Prediction with Chemical Language Models
Kwun Sy Lee, Jiawei Chen, Fuk Sheng Ford Chung +3
Reliable in silico molecular toxicity prediction is a cornerstone of modern drug discovery, offering a scalable alternative to experimental screening. However, the black-box nature…
cs.LG2025
Towards Robust Alignment of Language Models: Distributionally Robustifying Direct Preference Optimization
Junkang Wu, Yuexiang Xie, Zhengyi Yang +6
This study addresses the challenge of noise in training datasets for Direct Preference Optimization (DPO), a method for aligning Large Language Models (LLMs) with human preferences…