2 citations · 2 across the 4 of their papers we have counts for
4 papers
LM4LV: A Frozen Large Language Model for Low-level Vision Tasks
Boyang Zheng, Jinjin Gu, Shijun Li +1
The success of large language models (LLMs) has fostered a new research trend of multi-modality large language models (MLLMs), which changes the paradigm of various fields in compu…
DSDRNet: Disentangling Representation and Reconstruct Network for Domain Generalization
Juncheng Yang, Zuchao Li, Shuai Xie +2
Domain generalization faces challenges due to the distribution shift between training and testing sets, and the presence of unseen target domains. Common solutions include domain a…
Cross-Modal Adapter: Parameter-Efficient Transfer Learning Approach for Vision-Language Models
Juncheng Yang, Zuchao Li, Shuai Xie +3
Adapter-based parameter-efficient transfer learning has achieved exciting results in vision-language models. Traditional adapter methods often require training or fine-tuning, faci…
Soft-Prompting with Graph-of-Thought for Multi-modal Representation Learning
Juncheng Yang, Zuchao Li, Shuai Xie +3
The chain-of-thought technique has been received well in multi-modal tasks. It is a step-by-step linear reasoning process that adjusts the length of the chain to improve the perfor…