2 papers
cs.AI2026
FinRiskAtlas: Decision-Aligned Evaluation of Large Language Models for Financial Risk Review
Suyang Zhong, Jingzhe Zhu, Qi Xu +5
Deploying large language models for professional financial review requires more than measuring general financial competence: models must perform the specific review operation requi…
cs.MA2026
Organizational Control Layer: Governance Infrastructure at the Execution Boundary of LLM Agent Systems
Tianyu Shi, Yang Mo, Yiou Liu +6
LLM-based agents are increasingly deployed in workflows where generated outputs may trigger state-changing actions, such as price offers, refunds, payments, or tool calls. This cre…