3 papers
cs.CV2025
Advancing Video Self-Supervised Learning via Image Foundation Models
Jingwei Wu, Zhewei Huang, Chang Liu
In the past decade, image foundation models (IFMs) have achieved unprecedented progress. However, the potential of directly using IFMs for video self-supervised representation lear…
cs.CV2025
ViStoryBench: Comprehensive Benchmark Suite for Story Visualization
Cailin Zhuang, Ailin Huang, Yaoqi Hu +12
Story visualization aims to generate coherent image sequences that faithfully represent a narrative and match given character references. Despite progress in generative models, exi…
cs.CV2024
ARCON: Advancing Auto-Regressive Continuation for Driving Videos
Ruibo Ming, Jingwei Wu, Zhewei Huang +4
Recent advancements in auto-regressive large language models (LLMs) have led to their application in video generation. This paper explores the use of Large Vision Models (LVMs) for…