8 citations · 8 across the 1 of their papers we have counts for
1 paper
Yufeng Zhong, Long Xu, Jiebo Luo +1
3D dense captioning, as an emerging vision-language task, aims to identify and locate each object from a set of point clouds and generate a distinctive natural language sentence fo…