1 paper · 1 filter
Wei-Yao Wang, Kazuya Tateishi, Shuyang Cui +4
Multimodal representation learning has been shifting from traditional two-tower architectures to large language model (LLM)-based embedders due to their strong instruction-followin…