1 citations · 1 across the 1 of their papers we have counts for
3 papers
cs.CL2025
Automating Steering for Safe Multimodal Large Language Models
Lyucheng Wu, Mengru Wang, Ziwen Xu +4
Recent progress in Multimodal Large Language Models (MLLMs) has unlocked powerful cross-modal reasoning abilities, but also raised new safety concerns, particularly when faced with…
cs.CL2025
ReLearn: Unlearning via Learning for Large Language Models
Haoming Xu, Ningyuan Zhao, Liming Yang +7
Current unlearning methods for large language models usually rely on reverse optimization to reduce target token probabilities. However, this paradigm disrupts the subsequent token…
cs.CR2024★ 1 cited
PhishIntel: Toward Practical Deployment of Reference-Based Phishing Detection
Yuexin Li, Hiok Kuek Tan, Qiaoran Meng +6
Phishing is a critical cyber threat, exploiting deceptive tactics to compromise victims and cause significant financial losses. While reference-based phishing detectors (RBPDs) hav…