1 paper
Nesreen K. Ahmed, Nima Nafisi
Monitoring autonomous large language model (LLM) agents for covert malicious behavior is challenging due to delayed, context-dependent, and long-horizon attack patterns. Agents may…