2 papers
cs.CV2025
ViRectify: A Challenging Benchmark for Video Reasoning Correction with Multimodal Large Language Models
Xusen Hei, Jiali Chen, Jinyu Yang +2
As multimodal large language models (MLLMs) frequently exhibit errors in complex video reasoning scenarios, correcting these errors is critical for uncovering their weaknesses and…
cs.CV2025
ExpStar: Towards Automatic Commentary Generation for Multi-discipline Scientific Experiments
Jiali Chen, Yujie Jia, Zihan Wu +6
Experiment commentary is crucial in describing the experimental procedures, delving into underlying scientific principles, and incorporating content-related safety guidelines. In p…