7 citations · 9 across the 5 of their papers we have counts for
5 papers
Semi-Supervised Panoptic Narrative Grounding
Danni Yang, Jiayi Ji, Xiaoshuai Sun +4
Despite considerable progress, the advancement of Panoptic Narrative Grounding (PNG) remains hindered by costly annotations. In this paper, we introduce a novel Semi-Supervised Pan…
NICE: Improving Panoptic Narrative Detection and Segmentation with Cascading Collaborative Learning
Haowei Wang, Jiayi Ji, Tianyu Guo +4
Panoptic Narrative Detection (PND) and Segmentation (PNS) are two challenging tasks that involve identifying and locating multiple targets in an image according to a long narrative…
3D-STMN: Dependency-Driven Superpoint-Text Matching Network for End-to-End 3D Referring Expression Segmentation
Changli Wu, Yiwei Ma, Qi Chen +4
In 3D Referring Expression Segmentation (3D-RES), the earlier approach adopts a two-stage paradigm, extracting segmentation proposals and then matching them with referring expressi…
Towards Local Visual Modeling for Image Captioning
Yiwei Ma, Jiayi Ji, Xiaoshuai Sun +2
In this paper, we study the local visual modeling with grid features for image captioning, which is critical for generating accurate and detailed captions. To achieve this target,…
Towards Real-Time Panoptic Narrative Grounding by an End-to-End Grounding Network
Haowei Wang, Jiayi Ji, Yiyi Zhou +2
Panoptic Narrative Grounding (PNG) is an emerging cross-modal grounding task, which locates the target regions of an image corresponding to the text description. Existing approache…