2 papers
cs.CL2026
Goal Hijacking Attack on Large Language Models via Pseudo-Conversation Injection
Zheng Chen, Buhui Yao
Goal hijacking is a type of adversarial attack on Large Language Models (LLMs) where the objective is to manipulate the model into producing a specific, predetermined output, regar…
cs.SE2025
A Survey of Reinforcement Learning for Software Engineering
Dong Wang, Hanmo You, Lingwei Zhu +6
Reinforcement Learning (RL) has emerged as a powerful paradigm for sequential decision-making and has attracted growing interest across various domains, particularly following the…