98 citations · 158 across the 4 of their papers we have counts for
12 papers · 1 filter
End-to-End Multi-View Fusion for 3D Object Detection in LiDAR Point Clouds
Yin Zhou, Pei Sun, Yu Zhang +6
Recent work on 3D object detection advocates point cloud voxelization in birds-eye view, where objects preserve their physical dimensions and are naturally separable. When represen…
NOTE-RCNN: NOise Tolerant Ensemble RCNN for Semi-Supervised Object Detection
JIyang Gao, Jiang Wang, Shengyang Dai +2
The labeling cost of large number of bounding boxes is one of the main challenges for training modern object detectors. To reduce the dependence on expensive bounding box annotatio…
MAC: Mining Activity Concepts for Language-based Temporal Localization
Runzhou Ge, Jiyang Gao, Kan Chen +1
We address the problem of language-based temporal localization in untrimmed videos. Compared to temporal localization with fixed categories, this problem is more challenging as the…
CTAP: Complementary Temporal Action Proposal Generation
Jiyang Gao, Kan Chen, Ram Nevatia
Temporal action proposal generation is an important task, akin to object proposals, temporal action proposals are intended to capture "clips" or temporal intervals in videos that a…
Revisiting Temporal Modeling for Video-based Person ReID
Jiyang Gao, Ram Nevatia
Video-based person reID is an important task, which has received much attention in recent years due to the increasing demand in surveillance and camera networks. A typical video-ba…
Motion-Appearance Co-Memory Networks for Video Question Answering
Jiyang Gao, Runzhou Ge, Kan Chen +1
Video Question Answering (QA) is an important task in understanding video temporal structure. We observe that there are three unique attributes of video QA compared with image QA:…