3 papers
cs.CR2026
aiXamine: Unified Black-Box Evaluation of Cross-Dimensional Trade-offs in LLM Safety, Security, and Privacy
Fatih Deniz, Yazan Boshmaf, Dorde Popovic +1
The critical failure modes in deployed large language models (LLMs) are cross-dimensional: a model can score 99.3 in safety alignment while refusing one in three benign queries, or…
cs.CR2025
aiXamine: Simplified LLM Safety and Security
Fatih Deniz, Dorde Popovic, Yazan Boshmaf +4
Evaluating Large Language Models (LLMs) for safety and security remains a complex task, often requiring users to navigate a fragmented landscape of ad hoc benchmarks, datasets, met…
cs.CR2025
MANTIS: Detection of Zero-Day Malicious Domains Leveraging Low Reputed Hosting Infrastructure
Fatih Deniz, Mohamed Nabeel, Ting Yu +1
Internet miscreants increasingly utilize short-lived disposable domains to launch various attacks. Existing detection mechanisms are either too late to catch such malicious domains…