2 papers
cs.CL2026
Multiple-Debias: A Full-process Debiasing Method for Multilingual Pre-trained Language Models
Haoyu Liang, Peijian Zeng, Wentao Huang +2
Multilingual Pre-trained Language Models (MPLMs) have become essential tools for natural language processing. However, they often exhibit biases related to sensitive attributes suc…
cs.CL2025
Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models
Haoyu Liang, Youran Sun, Yunfeng Cai +2
The security issue of large language models (LLMs) has gained wide attention recently, with various defense mechanisms developed to prevent harmful output, among which safeguards b…