5 papers
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents
Abhishek Kumar, Carsten Maple
Large language models are increasingly deployed as IDE-integrated coding agents that decompose tasks, generate and edit files, run code, and refine outputs over many turns. Yet the…
PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines
Riya Tapwal, Abhishek Kumar, Carsten Maple
Multi-agent LLM systems introduce a security risk in which sensitive information accessed by one agent can propagate through shared context and reappear in downstream outputs, even…
Field-Localized Forgery Detection for Digital Identity Documents
Abhishek Kumar, Riya Tapwal, Carsten Maple +1
Digital onboarding and eKYC systems used by banks, fintech platforms, telecom providers, and other third-party services commonly verify users by comparing an uploaded identity docu…
Single-Configuration Attack Success Rate Is Not Enough: Jailbreak Evaluations Should Report Distributional Attack Success
Carsten Maple, Abhishek Kumar, Riya Tapwal
Many jailbreak attack research papers report attack success rates for a limited number of parameter settings, even though there are many combinations of parameter settings that cou…
Strengthening the No-Go Theorem for QRNGs
Vardaan Mongia, Abhishek Kumar, Shashi Prabhakar +1
Quantum random numbers are essential for security against quantum algorithms. Randomness as a beacon is a service being provided for companies and governments to upgrade their secu…