collaborators

7 papers

cs.CL2026

FinAuditing: A Financial Taxonomy-Structured Multi-Document Benchmark for Evaluating LLMs

Yan Wang, Keyi Wang, Shanshan Yang +12

Going beyond simple text processing, financial auditing requires detecting semantic, structural, and numerical inconsistencies across large-scale disclosures. As financial reports…

cs.CL2026

FinTagging: Benchmarking LLMs for Extracting and Structuring Financial Information

Yan Wang, Lingfei Qian, Xueqing Peng +18

Accurate interpretation of numerical data in financial reports is critical for markets and regulators. Although XBRL (eXtensible Business Reporting Language) provides a standard fo…

cs.CE2026

Evaluation and Benchmarking Suite for Financial Large Language Models and Agents

Shengyuan Lin, Kaiwen He, Jaisal Patel +10

Over the past three years, the financial services industry has witnessed Large Language Models (LLMs) and agents transitioning from the exploration stage to readiness and governanc…

cs.CE2026

zkFinGPT: Zero-Knowledge Proofs for Financial Generative Pre-trained Transformers

Xiao-Yang Liu, Ningjie Li, Keyi Wang +2

Financial Generative Pre-trained Transformers (FinGPT) with multimodal capabilities are now being increasingly adopted in various financial applications. However, due to the intell…

cs.AI2025

Reasoning Models Ace the CFA Exams

Jaisal Patel, Yunzhe Chen, Kaiwen He +4

Previous research has reported that large language models (LLMs) demonstrate poor performance on the Chartered Financial Analyst (CFA) exams. However, recent reasoning models have…

cs.CL2025

MultiFinBen: Benchmarking Large Language Models for Multilingual and Multimodal Financial Application

Xueqing Peng, Lingfei Qian, Yan Wang +44

Real-world financial analysis involves information across multiple languages and modalities, from reports and news to scanned filings and meeting recordings. Yet most existing eval…