2 papers
cs.CR2026
Automated jailbreak attack targeting multiple defense strategies
Qi Wang, Chengcheng Wan, Weijia He +4
Large language models (LLMs) have demonstrated remarkable capabilities across a wide range of tasks. However, their safety remains a critical concern due to their susceptibility to…
cs.AI2026
MagicAgent: Towards Generalized Agent Planning
Xuhui Ren, Shaokang Dong, Chen Yang +21
The evolution of Large Language Models (LLMs) from passive text processors to autonomous agents has established planning as a core component of modern intelligence. However, achiev…