1 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 1 cited
Speak Out of Turn: Safety Vulnerability of Large Language Models in Multi-turn Dialogue
Zhenhong Zhou, Jiuyang Xiang, Haopeng Chen +3
Large Language Models (LLMs) have been demonstrated to generate illegal or unethical responses, particularly when subjected to "jailbreak." Research on jailbreak has highlighted th…
cs.CL2023★ 1 cited
Quantifying and Analyzing Entity-level Memorization in Large Language Models
Zhenhong Zhou, Jiuyang Xiang, Chaomeng Chen +1
Large language models (LLMs) have been proven capable of memorizing their training data, which can be extracted through specifically designed prompts. As the scale of datasets cont…