Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
Hongyu Li, Songhao Han, Yue Liao +4
Understanding real-world videos with complex semantics and long temporal dependencies remains a fundamental challenge in computer vision. Recent progress in multimodal large langua…
cs.CV2024
Data Augmentation in Human-Centric Vision
Wentao Jiang, Yige Zhang, Shaozhong Zheng +2
This survey presents a comprehensive analysis of data augmentation techniques in human-centric vision tasks, a first of its kind in the field. It delves into a wide range of resear…