7 citations · 7 across the 5 of their papers we have counts for
Showing cs.CRShow all
2 papers · 1 filter
cs.CR2024★ 1 cited
Self and Cross-Model Distillation for LLMs: Effective Methods for Refusal Pattern Alignment
Jie Li, Yi Liu, Chongyang Liu +4
Large Language Models (LLMs) like OpenAI's GPT series, Anthropic's Claude, and Meta's LLaMa have shown remarkable capabilities in text generation. However, their susceptibility to…
cs.CR2024★ 7 cited
A Cross-Language Investigation into Jailbreak Attacks in Large Language Models
Jie Li, Yi Liu, Chongyang Liu +5
Large Language Models (LLMs) have become increasingly popular for their advanced text generation capabilities across various domains. However, like any software, they face security…