3 papers
cs.CV2024
DSDRNet: Disentangling Representation and Reconstruct Network for Domain Generalization
Juncheng Yang, Zuchao Li, Shuai Xie +2
Domain generalization faces challenges due to the distribution shift between training and testing sets, and the presence of unseen target domains. Common solutions include domain a…
cs.CV2024
Cross-Modal Adapter: Parameter-Efficient Transfer Learning Approach for Vision-Language Models
Juncheng Yang, Zuchao Li, Shuai Xie +3
Adapter-based parameter-efficient transfer learning has achieved exciting results in vision-language models. Traditional adapter methods often require training or fine-tuning, faci…
cs.AI2024
Soft-Prompting with Graph-of-Thought for Multi-modal Representation Learning
Juncheng Yang, Zuchao Li, Shuai Xie +3
The chain-of-thought technique has been received well in multi-modal tasks. It is a step-by-step linear reasoning process that adjusts the length of the chain to improve the perfor…