1 paper
Cong Wei, Yujie Zhong, Haoxian Tan +4
Boosted by Multi-modal Large Language Models (MLLMs), text-guided universal segmentation models for the image and video domains have made rapid progress recently. However, these me…