2 papers
cs.AI2026
More Accurate or More Efficient? Evaluating Locally Deployed Compact Open-Weight Language Models for Mathematical Reasoning
Orion Powers, Daniella Seum, Khaled Slhoub
Large language models are increasingly deployed on local hardware for privacy, cost, and accessibility reasons. Yet many evaluations emphasize accuracy while fewer quantify local r…
cs.CR2026
Can Open-Source LLM Agents Replace Static Application Security Testing Tools? An Empirical Assessment
Derek Yohn, Luke Flancher, Mirajul Islam +1
This paper explores the value of agentic AI tools for cybersecurity purposes. We evaluate the efficacy of a general-purpose GenAI Large Language Model- (GenAI-) based agent when po…