2 papers
cs.CL2025
Can We Reliably Rank Model Performance across Domains without Labeled Data?
Veronica Rammouz, Aaron Gonzalez, Carlos Cruzportillo +3
Estimating model performance without labels is an important goal for understanding how NLP models generalize. While prior work has proposed measures based on dataset similarity or…
cs.CR2025
Lateral Phishing With Large Language Models: A Large Organization Comparative Study
Mazal Bethany, Athanasios Galiopoulos, Emet Bethany +4
The emergence of Large Language Models (LLMs) has heightened the threat of phishing emails by enabling the generation of highly targeted, personalized, and automated attacks. Tradi…