Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality
Mohamed Elmoghany, Ryan Rossi, Seunghyun Yoon +26
Despite the significant progress that has been made in video generative models, existing state-of-the-art methods can only produce videos lasting 5-16 seconds, often labeled "long-…
cs.CV2024
Personalized Multimodal Large Language Models: A Survey
Junda Wu, Hanjia Lyu, Yu Xia +24
Multimodal Large Language Models (MLLMs) have become increasingly important due to their state-of-the-art performance and ability to integrate multiple data modalities, such as tex…