2 papers
cs.CL2024
Towards Tracing Trustworthiness Dynamics: Revisiting Pre-training Period of Large Language Models
Chen Qian, Jie Zhang, Wei Yao +5
Ensuring the trustworthiness of large language models (LLMs) is crucial. Most studies concentrate on fully pre-trained LLMs to better understand and improve LLMs' trustworthiness.…
cs.LG2024
Understanding Fairness Surrogate Functions in Algorithmic Fairness
Wei Yao, Zhanke Zhou, Zhicong Li +2
It has been observed that machine learning algorithms exhibit biased predictions against certain population groups. To mitigate such bias while achieving comparable accuracy, a pro…