2 papers
cs.AI2026
CIRCLE: A Framework for Evaluating AI from a Real-World Lens
Reva Schwartz, Carina Westling, Morgan Briggs +12
This paper proposes CIRCLE, a six-stage, lifecycle-based framework to bridge the reality gap between model-centric performance metrics and AI's materialized outcomes in deployment.…
cs.CR2026
Eliciting Least-to-Most Reasoning for Phishing URL Detection
Holly Trikilis, Pasindu Marasinghe, Fariza Rashid +1
Phishing continues to be one of the most prevalent attack vectors, making accurate classification of phishing URLs essential. Recently, large language models (LLMs) have demonstrat…