93 citations · 112 across the 7 of their papers we have counts for
4 papers · 1 filter
Hierarchical Point Cloud Encoding and Decoding with Lightweight Self-Attention based Model
En Yen Puang, Hao Zhang, Hongyuan Zhu +1
In this paper we present SA-CNN, a hierarchical and lightweight self-attention based encoding and decoding architecture for representation learning of point cloud data. The propose…
Towards Debiasing Temporal Sentence Grounding in Video
Hao Zhang, Aixin Sun, Wei Jing +1
The temporal sentence grounding in video (TSGV) task is to locate a temporal moment from an untrimmed video, to match a language query, i.e., a sentence. Without considering bias i…
Domain Generalization for Vision-based Driving Trajectory Generation
Yunkai Wang, Dongkun Zhang, Yuxiang Cui +5
One of the challenges in vision-based driving trajectory generation is dealing with out-of-distribution scenarios. In this paper, we propose a domain generalization method for visi…
6D Pose Estimation with Correlation Fusion
Yi Cheng, Hongyuan Zhu, Ying Sun +6
6D object pose estimation is widely applied in robotic tasks such as grasping and manipulation. Prior methods using RGB-only images are vulnerable to heavy occlusion and poor illum…