2 papers
cs.CR2026
Do Defenses Against LLM Extraction Work Across Attacks? A Lifecycle Benchmark of Black-Box Model Extraction
Shuze Liu, Kaixiang Zhao, Runyang Xu +4
Large language models (LLMs) deployed through text-only APIs face model extraction risks, as adversaries can collect their responses to train surrogates that reproduce their capabi…
cs.CR2026
APTInvestBench: Evaluating Autonomous APT Investigation under Varying Telemetry
Yu Wang, Shuhao Li, Tao Yin +4
Large language model (LLM) agents could help security operations centers (SOCs) investigate advanced persistent threats (APTs) by turning weak leads into evidence for intrusion sco…