1 citations · 1 across the 1 of their papers we have counts for
1 paper
Zhongwei Ren, Zhicheng Huang, Yunchao Wei +4
While large multimodal models (LMMs) have achieved remarkable progress, generating pixel-level masks for image reasoning tasks involving multiple open-world targets remains a chall…