From the 1 of 4 linked papers with an AI index.
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Reflect to Inform: Boosting Multimodal Reasoning via Information-Gain-Driven Verification
Shuai Lv, Chang Liu, Feng Tang +5
Multimodal Large Language Models (MLLMs) achieve strong multimodal reasoning performance, yet we identify a recurring failure mode in long-form generation: as outputs grow longer,…
cs.CV2025
T2VEval: Benchmark Dataset and Objective Evaluation Method for T2V-generated Videos
Zelu Qi, Ping Shi, Shuqi Wang +7
Recent advances in text-to-video (T2V) technology, as demonstrated by models such as Runway Gen-3, Pika, Sora, and Kling, have significantly broadened the applicability and popular…