3 papers
cs.CL2025
Detection, Classification, and Mitigation of Gender Bias in Large Language Models
Xiaoqing Cheng, Hongying Zan, Lulu Kong +2
With the rapid development of large language models (LLMs), they have significantly improved efficiency across a wide range of domains. However, recent studies have revealed that L…
cs.CL2025
BiasFilter: An Inference-Time Debiasing Framework for Large Language Models
Xiaoqing Cheng, Ruizhe Chen, Hongying Zan +2
Mitigating social bias in large language models (LLMs) has become an increasingly important research objective. However, existing debiasing methods often incur high human and compu…
cs.CL2024
OpenEval: Benchmarking Chinese LLMs across Capability, Alignment and Safety
Chuang Liu, Linhao Yu, Jiaxuan Li +11
The rapid development of Chinese large language models (LLMs) poses big challenges for efficient LLM evaluation. While current initiatives have introduced new benchmarks or evaluat…