2 papers
cs.CL2025
SCOP: Evaluating the Comprehension Process of Large Language Models from a Cognitive View
Yongjie Xiao, Hongru Liang, Peixin Qin +2
Despite the great potential of large language models(LLMs) in machine comprehension, it is still disturbing to fully count on them in real-world scenarios. This is probably because…
cs.CL2025
VisualSimpleQA: A Benchmark for Decoupled Evaluation of Large Vision-Language Models in Fact-Seeking Question Answering
Yanling Wang, Yihan Zhao, Xiaodong Chen +7
Large vision-language models (LVLMs) have demonstrated remarkable achievements, yet the generation of non-factual responses remains prevalent in fact-seeking question answering (QA…