benchmark dataset 1contextual relevance 1cross-modal evaluation 1large language models 1scientific figure quality assessment 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.CV2026
SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context
Zihan Deng, Chuanzhi Xu, Huiqi Liang +3
The paper introduces SciFigQual-Bench, a benchmark dataset that evaluates the quality of scientific figures within full manuscript context across five dimensions, and presents a cr…
cs.CL2026
COTCAgent: Preventive Consultation via Probabilistic Chain-of-Thought Completion
Zihan Deng, Xiaozhen Zhong, Chuanzhi Xu
As large language models empower healthcare, intelligent clinical decision support has developed rapidly. Longitudinal electronic health records (EHR) provide essential temporal ev…