4 citations · 4 across the 4 of their papers we have counts for
1 paper · 2 filters
Junfei Xiao, Zheng Xu, Alan Yuille +2
This paper demonstrates that a progressively aligned language model can effectively bridge frozen vision encoders and large language models (LLMs). While the fundamental architectu…