34 citations · 95 across the 17 of their papers we have counts for
1 paper · 2 filters
Chengxi Zeng, Yuxuan Jiang, Ge Gao +6
Vision-language segmentation models such as SAM3 enable flexible, prompt-driven visual grounding, but inherit large, general-purpose text encoders originally designed for open-ende…