1 paper
Renjie Liang, Zijian Xu
Building a 3D CT vision language model begins with a choice of which image encoder to build on. Today that choice is made by fine-tuning every candidate through the full language m…