5 papers
Multimodal Continual Instruction Tuning with Dynamic Gradient Guidance
Songze Li, Mingyu Gao, Tonghua Su +2
Multimodal continual instruction tuning enables multimodal large language models to sequentially adapt to new tasks while building upon previously acquired knowledge. However, this…
Global-Local Aware Scene Text Editing
Fuxiang Yang, Tonghua Su, Donglin Di +4
Scene Text Editing (STE) involves replacing text in a scene image with new target text while preserving both the original text style and background texture. Existing methods suffer…
GUSLO: General and Unified Structured Light Optimization
Tinglei Wan, Tonghua Su, Zhongjie Wang
Structured light (SL) 3D reconstruction captures the precise surface shape of objects, providing high-accuracy 3D data essential for industrial inspection and cultural heritage dig…
Balancing Stability and Plasticity in Pretrained Detector: A Dual-Path Framework for Incremental Object Detection
Songze Li, Qixing Xu, Tonghua Su +2
The balance between stability and plasticity remains a fundamental challenge in pretrained model-based incremental object detection (PTMIOD). While existing PTMIOD methods demonstr…
DUKAE: DUal-level Knowledge Accumulation and Ensemble for Pre-Trained Model-Based Continual Learning
Songze Li, Tonghua Su, Xu-Yao Zhang +2
Pre-trained model-based continual learning (PTMCL) has garnered growing attention, as it enables more rapid acquisition of new knowledge by leveraging the extensive foundational un…