2 papers
cs.CE2026
AuditFraudBench: Benchmarking Audit Judgment in Detecting Fraudulent Misstatements
Zhiwei Liu, Yueru He, Qing Ou +4
Large language models (LLMs) have shown strong performance in financial analysis and surface-level factual error detection, yet their ability to identify fraudulent financial misin…
cs.CE2026
AutoRedTrader: Autonomous Red Teaming of Trading Agents through Synthetic Misinformation Injection
Zhiwei Liu, Yangyang Yu, Yupeng Cao +8
LLM-based financial agents increasingly rely on both numerical market data and textual signals for sequential trading and stock prediction. However, financial misinformation often…