18 citations · 25 across the 3 of their papers we have counts for
6 papers
LocATe: End-to-end Localization of Actions in 3D with Transformers
Jiankai Sun, Bolei Zhou, Michael J. Black +1
Understanding a person's behavior from their 3D motion is a fundamental problem in computer vision with many applications. An important component of this problem is 3D Temporal Act…
Transferable Active Grasping and Real Embodied Dataset
Xiangyu Chen, Zelin Ye, Jiankai Sun +4
Grasping in cluttered scenes is challenging for robot vision systems, as detection accuracy can be hindered by partial occlusion of objects. We adopt a reinforcement learning (RL)…
SegVoxelNet: Exploring Semantic Context and Depth-aware Features for 3D Vehicle Detection from Point Cloud
Hongwei Yi, Shaoshuai Shi, Mingyu Ding +6
3D vehicle detection based on point cloud is a challenging task in real-world applications such as autonomous driving. Despite significant progress has been made, we observe two as…
GDRQ: Group-based Distribution Reshaping for Quantization
Haibao Yu, Tuopu Wen, Guangliang Cheng +3
Low-bit quantization is challenging to maintain high performance with limited model capacity (e.g., 4-bit for both weights and activations). Naturally, the distribution of both wei…
Cross-view Semantic Segmentation for Sensing Surroundings
Bowen Pan, Jiankai Sun, Ho Yin Tiga Leung +2
Sensing surroundings plays a crucial role in human spatial perception, as it extracts the spatial configuration of objects as well as the free space from the observations. To facil…
NavigationNet: A Large-scale Interactive Indoor Navigation Dataset
He Huang, Yujing Shen, Jiankai Sun +1
Indoor navigation aims at performing navigation within buildings. In scenes like home and factory, most intelligent mobile devices require an functionality of routing to guide itse…