3 citations · 3 across the 4 of their papers we have counts for
Showing cs.CRShow all
2 papers · 1 filter
cs.CR2025
Black-box Optimization of LLM Outputs by Asking for Directions
Jie Zhang, Meng Ding, Yang Liu +2
We present a novel approach for attacking black-box large language models (LLMs) by exploiting their ability to express confidence in natural language. Existing black-box attacks r…
cs.CR2025
AgentArmor: Enforcing Program Analysis on Agent Runtime Trace to Defend Against Prompt Injection
Peiran Wang, Yang Liu, Yunfei Lu +6
Large Language Model (LLM) agents offer a powerful new paradigm for solving various problems by combining natural language reasoning with the execution of external tools. However,…