2 citations · 2 across the 6 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
FinRiskAtlas: Decision-Aligned Evaluation of Large Language Models for Financial Risk Review
Suyang Zhong, Jingzhe Zhu, Qi Xu +5
Deploying large language models for professional financial review requires more than measuring general financial competence: models must perform the specific review operation requi…
cs.AI2026
SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images
Jialu Shen, Han Lyu, Suyang Zhong +5
Spectra are a prevalent yet highly information-dense form of scientific imagery, presenting substantial challenges to multimodal large language models (MLLMs) due to their unstruct…