papers
Publications (3)
cs.CL2026
APEX-Accounting
Julien Benchek, Austin Bennett, Jasmin Kern +8
The paper presents APEX-Accounting, a benchmark for evaluating how well advanced language models can perform real accounting tasks such as reconciliation, expense accrual, transact…
#accounting automation#benchmark#large language models#evaluation metrics
cs.CR2024
Facade: High-Precision Insider Threat Detection Using Deep Contextual Anomaly Detection
Alex Kantchelian, Casper Neo, Ryan Stevens +10
Insiders with privileged access have the power to cause great harm to their organization. Even a single insider threat incident can be catastrophic, resulting in both financial los…
econ.GN2026
Payrolls to Prompts: Firm-Level Evidence on the Substitution of Labor for AI
Ryan Stevens
Generative AI has the potential to transform how firms produce output. Yet, credible evidence on how AI is actually substituting for human labor remains limited. In this paper, we…