1 paper · 1 filter
Dongchao Yang, Haohan Guo, Yuanyuan Wang +5
The Large Language models (LLMs) have demonstrated supreme capabilities in text understanding and generation, but cannot be directly applied to cross-modal tasks without fine-tunin…