4 citations · 8 across the 5 of their papers we have counts for
5 papers
Towards Improved Preference Optimization Pipeline: from Data Generation to Budget-Controlled Regularization
Zhuotong Chen, Fang Liu, Jennifer Zhu +2
Direct Preference Optimization (DPO) and its variants have become the de facto standards for aligning large language models (LLMs) with human preferences or specific goals. However…
PID Control-Based Self-Healing to Improve the Robustness of Large Language Models
Zhuotong Chen, Zihu Wang, Yifan Yang +2
Despite the effectiveness of deep neural networks in numerous natural language processing applications, recent findings have exposed the vulnerability of these language models when…
Logic-Scaffolding: Personalized Aspect-Instructed Recommendation Explanation Generation using LLMs
Behnam Rahdari, Hao Ding, Ziwei Fan +4
The unique capabilities of Large Language Models (LLMs), such as the natural language text generation ability, position them as strong candidates for providing explanation for reco…
Asymptotically Fair Participation in Machine Learning Models: an Optimal Control Perspective
Zhuotong Chen, Qianxiao Li, Zheng Zhang
The performance of state-of-the-art machine learning models often deteriorates when testing on demographics that are under-represented in the training dataset. This problem has pre…
Self-Healing Robust Neural Networks via Closed-Loop Control
Zhuotong Chen, Qianxiao Li, Zheng Zhang
Despite the wide applications of neural networks, there have been increasing concerns about their vulnerability issue. While numerous attack and defense techniques have been develo…