19 citations · 28 across the 6 of their papers we have counts for
4 papers · 1 filter
One Jailbreak, Many Tongues: Learning Language-Insensitive Intention Representations for Multilingual Jailbreak Detection
Shuyu Jiang, Kaiyu Xu, Xingshu Chen +5
Large language models (LLMs) are increasingly deployed in applications for global multilingual users, yet safety training remains concentrated in dominant languages and has not pro…
Prompt Packer: Deceiving LLMs through Compositional Instruction with Hidden Attacks
Shuyu Jiang, Xingshu Chen, Rui Tang
Recently, Large language models (LLMs) with powerful general capabilities have been increasingly integrated into various Web applications, while undergoing alignment training to en…
ReZG: Retrieval-Augmented Zero-Shot Counter Narrative Generation for Hate Speech
Shuyu Jiang, Wenyi Tang, Xingshu Chen +3
The proliferation of hate speech (HS) on social media poses a serious threat to societal security. Automatic counter narrative (CN) generation, as an active strategy for HS interve…
ClueGraphSum: Let Key Clues Guide the Cross-Lingual Abstractive Summarization
Shuyu Jiang, Dengbiao Tu, Xingshu Chen +3
Cross-Lingual Summarization (CLS) is the task to generate a summary in one language for an article in a different language. Previous studies on CLS mainly take pipeline methods or…