Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Goal Hijacking Attack on Large Language Models via Pseudo-Conversation Injection
Zheng Chen, Buhui Yao
Goal hijacking is a type of adversarial attack on Large Language Models (LLMs) where the objective is to manipulate the model into producing a specific, predetermined output, regar…
cs.CL2023
KwaiYiiMath: Technical Report
Jiayi Fu, Lei Lin, Xiaoyang Gao +18
Recent advancements in large language models (LLMs) have demonstrated remarkable abilities in handling a variety of natural language processing (NLP) downstream tasks, even on math…