2 papers
cs.CL2026
A Systematic Evaluation of Positional Bias in Multi-Video Summarization with MLLMs
Huangchen Xu, Yuan Wu, Yi Chang
Multimodal Large Language Models (MLLMs) are increasingly used for video understanding, yet their reliability under multi-video inputs remains poorly understood. We study positiona…
cs.CL2026
VCIFBench: Evaluating Complex Instruction Following for Video Understanding
Huangchen Xu, Yuan Wu, Yi Chang
Multimodal large language models have made rapid progress in video understanding, yet existing benchmarks largely rely on simple prompts and provide limited evidence about whether…