2 papers
cs.CR2025
Active Honeypot Guardrail System: Probing and Confirming Multi-Turn LLM Jailbreaks
ChenYu Wu, Yi Wang, Yang Liao
Large language models (LLMs) are increasingly vulnerable to multi-turn jailbreak attacks, where adversaries iteratively elicit harmful behaviors that bypass single-turn safety filt…
cs.CR2025
BERTector: An Intrusion Detection Framework Constructed via Joint-dataset Learning Based on Language Model
Haoyang Hu, Xun Huang, Chenyu Wu +3
Intrusion detection systems (IDS) are widely used to maintain the stability of network environments, but still face restrictions in generalizability due to the heterogeneity of net…