3 papers
cs.CV2025
Bring Your Dreams to Life: Continual Text-to-Video Customization
Jiahua Dong, Xudong Wang, Wenqi Liang +7
Customized text-to-video generation (CTVG) has recently witnessed great progress in generating tailored videos from user-specific text. However, most CTVG methods assume that perso…
cs.CV2025
COLT: Enhancing Video Large Language Models with Continual Tool Usage
Yuyang Liu, Meng Cao, Xinyuan Shi +1
The success of Large Language Models (LLMs) has significantly propelled the research of video understanding. To harvest the benefits of well-trained expert models (i.e., tools), vi…
cs.CV2024
Continual LLaVA: Continual Instruction Tuning in Large Vision-Language Models
Meng Cao, Yuyang Liu, Yingfei Liu +6
Instruction tuning constitutes a prevalent technique for tailoring Large Vision Language Models (LVLMs) to meet individual task requirements. To date, most of the existing approach…