3 papers
cs.CE2026
Evaluation and Benchmarking Suite for Financial Large Language Models and Agents
Shengyuan Lin, Kaiwen He, Jaisal Patel +10
Over the past three years, the financial services industry has witnessed Large Language Models (LLMs) and agents transitioning from the exploration stage to readiness and governanc…
cs.AI2025
Reasoning Models Ace the CFA Exams
Jaisal Patel, Yunzhe Chen, Kaiwen He +4
Previous research has reported that large language models (LLMs) demonstrate poor performance on the Chartered Financial Analyst (CFA) exams. However, recent reasoning models have…
cs.AI2025
Recon-Act: A Self-Evolving Multi-Agent Browser-Use System via Web Reconnaissance, Tool Generation, and Task Execution
Kaiwen He, Zhiwei Wang, Chenyi Zhuang +1
Recent years, multimodal models have made remarkable strides and pave the way for intelligent browser use agents. However, when solving tasks on real world webpages in multi-turn,…