activity
20192022
most citedDistributed Attention for Grounded Image Captioning

19 citations · 55 across the 8 of their papers we have counts for

collaborators

10 papers

cs.CV2022

Self-Supervised Image Representation Learning with Geometric Set Consistency

Nenglun Chen, Lei Chu, Hao Pan +2

We propose a method for self-supervised image representation learning under the guidance of 3D geometric consistency. Our intuition is that 3D geometric consistency priors such as…

cs.CV20224 cited

Towards 3D Scene Understanding by Referring Synthetic Models

Runnan Chen, Xinge Zhu, Nenglun Chen +5

Promising performance has been achieved for visual perception on the point cloud. However, the current methods typically rely on labour-extensive annotations on the scene scans. In…

cs.CV2021

PR-Net: Preference Reasoning for Personalized Video Highlight Detection

Runnan Chen, Penghao Zhou, Wenzhe Wang +4

Personalized video highlight detection aims to shorten a long video to interesting moments according to a user's preference, which has recently raised the community's attention. Cu…

cs.CV202119 cited

Distributed Attention for Grounded Image Captioning

Nenglun Chen, Xingjia Pan, Runnan Chen +7

We study the problem of weakly supervised grounded image captioning. That is, given an image, the goal is to automatically generate a sentence describing the context of the image w…

cs.GR202114 cited

CurveFusion: Reconstructing Thin Structures from RGBD Sequences

Lingjie Liu, Nenglun Chen, Duygu Ceylan +3

We introduce CurveFusion, the first approach for high quality scanning of thin structures at interactive rates using a handheld RGBD camera. Thin filament-like structures are mathe…

cs.CV20203 cited

Point2Skeleton: Learning Skeletal Representations from Point Clouds

Cheng Lin, Changjian Li, Yuan Liu +3

We introduce Point2Skeleton, an unsupervised method to learn skeletal representations from point clouds. Existing skeletonization methods are limited to tubular shapes and the stri…