3 citations · 3 across the 1 of their papers we have counts for
3 papers
cs.CV2026★ 3 cited
High-Quality Entity Segmentation and Grounding
Lu Qi, Yi-Wen Chen, Tao Zhang +4
In this work, we propose ESG, a pipeline for high-quality entity segmentation and grounding supported by a new dataset EntitySeg. At first, the proposed dataset naming EntitySeg co…
cs.CV2025
Unified Dense Prediction of Video Diffusion
Lehan Yang, Lu Qi, Xiangtai Li +3
We present a unified network for simultaneously generating videos and their corresponding entity segmentation and depth maps from text prompts. We utilize colormap to represent ent…
cs.CV2025
VRMDiff: Text-Guided Video Referring Matting Generation of Diffusion
Lehan Yang, Jincen Song, Tianlong Wang +4
We propose a new task, video referring matting, which obtains the alpha matte of a specified instance by inputting a referring caption. We treat the dense prediction task of mattin…