Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
VL-RewardBench: A Challenging Benchmark for Vision-Language Generative Reward Models
Lei Li, Yuancheng Wei, Zhihui Xie +9
Vision-language generative reward models (VL-GenRMs) play a crucial role in aligning and evaluating multimodal AI systems, yet their own evaluation remains under-explored. Current…
cs.CV2024
Temporal Reasoning Transfer from Text to Video
Lei Li, Yuanxin Liu, Linli Yao +6
Video Large Language Models (Video LLMs) have shown promising capabilities in video comprehension, yet they struggle with tracking temporal changes and reasoning about temporal rel…