From the 1 of 3 linked papers with an AI index.
3 papers
cs.CL2026
HarDBench: A Benchmark for Draft-Based Co-Authoring Jailbreak Attacks for Safe Human-LLM Collaborative Writing
Euntae Kim, Soomin Han, Buru Chang
The paper introduces HarDBench, a benchmark that evaluates how vulnerable large language models are to jailbreak attacks when used as co-authors in draft-based writing, and propose…
cs.CR2026
KidnapRAG: A Black-Box Attack for Hijacking Reasoning in Agentic Retrieval-Augmented Generation Systems
Chanwoo Choi, Euntae Kim, Kyuho Lee +6
Retrieval-Augmented Generation (RAG) systems are vulnerable to poisoning attacks that inject malicious documents into the retrieval process to manipulate model outputs. Recent Agen…
cs.CV2025
NOAH: Benchmarking Narrative Prior driven Hallucination and Omission in Video Large Language Models
Kyuho Lee, Euntae Kim, Jinwoo Choi +1
Video large language models (Video LLMs) have recently achieved strong performance on tasks such as captioning, summarization, and question answering. Many models and training meth…