3 papers
cs.CL2026
Correcting Mean Bias in Text Embeddings: A Refined Renormalization with Training-Free Improvements on MMTEB
Xingyu Ren, Youran Sun, Haoyu Liang
We find that current sentence-embedding models produce outputs with a consistent bias: every embedding decomposes as , where the mean is near-identical acro…
cs.CL2026
Multiple-Debias: A Full-process Debiasing Method for Multilingual Pre-trained Language Models
Haoyu Liang, Peijian Zeng, Wentao Huang +2
Multilingual Pre-trained Language Models (MPLMs) have become essential tools for natural language processing. However, they often exhibit biases related to sensitive attributes suc…
cs.CL2025
Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models
Haoyu Liang, Youran Sun, Yunfeng Cai +2
The security issue of large language models (LLMs) has gained wide attention recently, with various defense mechanisms developed to prevent harmful output, among which safeguards b…