2 papers
cs.LG2026
Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents
Pengfei He, Ash Fox, Lesly Miculicich +7
Large language models (LLMs) have shown promise in assisting cybersecurity tasks, yet existing approaches struggle with automatic vulnerability discovery and exploitation due to li…
cs.LG2025
A Unifying Information-theoretic Perspective on Evaluating Generative Models
Alexis Fox, Samarth Swarup, Abhijin Adiga
Considering the difficulty of interpreting generative model output, there is significant current research focused on determining meaningful evaluation metrics. Several recent appro…