Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
One Jailbreak, Many Tongues: Learning Language-Insensitive Intention Representations for Multilingual Jailbreak Detection
Shuyu Jiang, Kaiyu Xu, Xingshu Chen +5
Large language models (LLMs) are increasingly deployed in applications for global multilingual users, yet safety training remains concentrated in dominant languages and has not pro…
cs.CL2026
UGID: Unified Graph Isomorphism for Debiasing Large Language Models
Zikang Ding, Junchi Yao, Junhao Li +4
Large language models (LLMs) exhibit pronounced social biases. Output-level or data-optimization--based debiasing methods cannot fully resolve these biases, and many prior works ha…
cs.CL2024
FuxiTranyu: A Multilingual Large Language Model Trained with Balanced Data
Haoran Sun, Renren Jin, Shaoyang Xu +10
Large language models (LLMs) have demonstrated prowess in a wide range of tasks. However, many LLMs exhibit significant performance discrepancies between high- and low-resource lan…