3 papers
cs.AI2026
Redundant or Necessary? A Benchmark for Detecting Redundant Steps in Agent Trajectories
Minyang Hu, Bo Yang, Zhinuo Zhou +4
LLM-based agents have demonstrated strong capabilities in solving complex tasks through multi-step reasoning and tool use. However, existing evaluation protocols primarily focus on…
cs.CR2025
CyLens: Towards Reinventing Cyber Threat Intelligence in the Paradigm of Agentic Large Language Models
Xiaoqun Liu, Jiacheng Liang, Qiben Yan +5
The exponential growth of cyber threat knowledge, exemplified by the expansion of databases such as MITRE-CVE and NVD, poses significant challenges for cyber threat analysis. Secur…
cs.CR2025
Data to Defense: The Role of Curation in Customizing LLMs Against Jailbreaking Attacks
Xiaoqun Liu, Jiacheng Liang, Luoxi Tang +3
Large language models (LLMs) are widely adapted for downstream applications through fine-tuning, a process named customization. However, recent studies have identified a vulnerabil…