2 papers
cs.CL2025
Golden Touchstone: A Comprehensive Bilingual Benchmark for Evaluating Financial Large Language Models
Xiaojun Wu, Junxi Liu, Huanyi Su +10
As large language models (LLMs) increasingly permeate the financial sector, there is a pressing need for a standardized method to comprehensively assess their performance. Existing…
q-fin.CP2025
QuantBench: Benchmarking AI Methods for Quantitative Investment
Saizhuo Wang, Hao Kong, Jiadong Guo +7
The field of artificial intelligence (AI) in quantitative investment has seen significant advancements, yet it lacks a standardized benchmark aligned with industry practices. This…