Showing 2024Show all
2 papers · 1 filter
cs.CR2024
"Moralized" Multi-Step Jailbreak Prompts: Black-Box Testing of Guardrails in Large Language Models for Verbal Attacks
Libo Wang
As the application of large language models continues to expand in various fields, it poses higher challenges to the effectiveness of identifying harmful content generation and gua…
cs.LG2024
Reducing Reasoning Costs: The Path of Optimization for Chain of Thought via Sparse Attention Mechanism
Libo Wang
In order to address the chain of thought in the large language model inference cost surge, this research proposes to use a sparse attention mechanism that only focuses on a few rel…