3 papers
cs.AI2026
Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations
Matteo Gioele Collu, Riccardo Conte, Alberto Giaretta +4
In this paper, we investigate whether refusal behavior can be predicted from LLM intermediate activations before decoding using linear probes trained on residual stream activations…
cs.CR2025
WeiDetect: Weibull Distribution-Based Defense against Poisoning Attacks in Federated Learning for Network Intrusion Detection Systems
Sameera K. M., Vinod P., Anderson Rocha +2
In the era of data expansion, ensuring data privacy has become increasingly critical, posing significant challenges to traditional AI-based applications. In addition, the increasin…
cs.IR2024
IntellBot: Retrieval Augmented LLM Chatbot for Cyber Threat Knowledge Delivery
Dincy R. Arikkat, Abhinav M., Navya Binu +6
In the rapidly evolving landscape of cyber security, intelligent chatbots are gaining prominence. Artificial Intelligence, Machine Learning, and Natural Language Processing empower…