12 citations · 13 across the 7 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.CL2024
Shaping the Safety Boundaries: Understanding and Defending Against Jailbreaks in Large Language Models
Lang Gao, Jiahui Geng, Xiangliang Zhang +2
Jailbreaking in Large Language Models (LLMs) is a major security concern as it can deceive LLMs to generate harmful text. Yet, there is still insufficient understanding of how jail…
cs.CL2024
Write Summary Step-by-Step: A Pilot Study of Stepwise Summarization
Xiuying Chen, Shen Gao, Mingzhe Li +3
Nowadays, neural text generation has made tremendous progress in abstractive summarization tasks. However, most of the existing summarization models take in the whole document all…