2 papers
cs.CR2026
KidnapRAG: A Black-Box Attack for Hijacking Reasoning in Agentic Retrieval-Augmented Generation Systems
Chanwoo Choi, Euntae Kim, Kyuho Lee +6
Retrieval-Augmented Generation (RAG) systems are vulnerable to poisoning attacks that inject malicious documents into the retrieval process to manipulate model outputs. Recent Agen…
cs.CL2025
Exploring the Impact of Instruction-Tuning on LLM's Susceptibility to Misinformation
Kyubeen Han, Junseo Jang, Hongjin Kim +2
Instruction-tuning enhances the ability of large language models (LLMs) to follow user instructions more accurately, improving usability while reducing harmful outputs. However, th…