Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
DECO: Unleashing the Potential of ConvNets for Query-based Detection and Segmentation
Xinghao Chen, Siwei Li, Yijing Yang +1
Transformer and its variants have shown great potential for various vision tasks in recent years, including image classification, object detection and segmentation. Meanwhile, rece…
cs.CV2024
Diagnosing and Re-learning for Balanced Multimodal Learning
Yake Wei, Siwei Li, Ruoxuan Feng +1
To overcome the imbalanced multimodal learning problem, where models prefer the training of specific modalities, existing methods propose to control the training of uni-modal encod…