Publications (15)
R4D: Utilizing Reference Objects for Long-Range Distance Estimation
Yingwei Li, Tiffany Chen, Maya Kabkab +4
Estimating the distance of objects is a safety-critical task for autonomous driving. Focusing on short-range objects, existing methods and datasets neglect the equally important lo…
NISP: Pruning Networks using Neuron Importance Score Propagation
Ruichi Yu, Ang Li, Chun-Fu Chen +6
To reduce the significant redundancy in deep Convolutional Neural Networks (CNNs), most existing methods prune neurons by only considering statistics of an individual layer or two…
Layout-induced Video Representation for Recognizing Agent-in-Place Actions
Ruichi Yu, Hongcheng Wang, Ang Li +3
We address the recognition of agent-in-place actions, which are associated with agents who perform them and places where they occur, in the context of outdoor home surveillance. We…
Visual Relationship Detection with Internal and External Linguistic Knowledge Distillation
Ruichi Yu, Ang Li, Vlad I. Morariu +1
Understanding visual relationships involves identifying the subject, the object, and a predicate relating them. We leverage the strong correlations between the predicate and the (s…
Uncertainty Modeling of Contextual-Connections between Tracklets for Unconstrained Video-based Face Recognition
Jingxiao Zheng, Ruichi Yu, Jun-Cheng Chen +3
Unconstrained video-based face recognition is a challenging problem due to significant within-video variations caused by pose, occlusion and blur. To tackle this problem, an effect…
The Role of Context Selection in Object Detection
Ruichi Yu, Xi Chen, Vlad I. Morariu +1
We investigate the reasons why context in object detection has limited utility by isolating and evaluating the predictive power of different context cues under ideal conditions in…
ReMotENet: Efficient Relevant Motion Event Detection for Large-scale Home Surveillance Videos
Ruichi Yu, Hongcheng Wang, Larry S. Davis
This paper addresses the problem of detecting relevant motion caused by objects of interest (e.g., person and vehicles) in large scale home surveillance videos. The traditional met…
Modeling Local Geometric Structure of 3D Point Clouds using Geo-CNN
Shiyi Lan, Ruichi Yu, Gang Yu +1
Recent advances in deep convolutional neural networks (CNNs) have motivated researchers to adapt CNNs to directly model points in 3D point clouds. Modeling local structure has been…
Dynamic Zoom-in Network for Fast Object Detection in Large Images
Mingfei Gao, Ruichi Yu, Ang Li +2
We introduce a generic framework that reduces the computational cost of object detection while retaining accuracy for scenarios where objects with varied sizes appear in high resol…
SoDA: Multi-Object Tracking with Soft Data Association
Wei-Chih Hung, Henrik Kretzschmar, Tsung-Yi Lin +4
Robust multi-object tracking (MOT) is a prerequisite fora safe deployment of self-driving cars. Tracking objects, however, remains a highly challenging problem, especially in clutt…
STT: Stateful Tracking with Transformers for Autonomous Driving
Longlong Jing, Ruichi Yu, Xu Chen +19
Tracking objects in three-dimensional space is critical for autonomous driving. To ensure safety while driving, the tracker must be able to reliably track objects across frames and…
Depth Estimation Matters Most: Improving Per-Object Depth Estimation for Monocular 3D Detection and Tracking
Longlong Jing, Ruichi Yu, Henrik Kretzschmar +11
Monocular image-based 3D perception has become an active research area in recent years owing to its applications in autonomous driving. Approaches to monocular 3D perception includ…
C-WSL: Count-guided Weakly Supervised Localization
Mingfei Gao, Ang Li, Ruichi Yu +2
We introduce count-guided weakly supervised localization (C-WSL), an approach that uses per-class object count as a new form of supervision to improve weakly supervised localizatio…
VITON: An Image-based Virtual Try-on Network
Xintong Han, Zuxuan Wu, Zhe Wu +2
We present an image-based VIirtual Try-On Network (VITON) without using 3D information in any form, which seamlessly transfers a desired clothing item onto the corresponding region…
Generating Holistic 3D Scene Abstractions for Text-based Image Retrieval
Ang Li, Jin Sun, Joe Yue-Hei Ng +3
Spatial relationships between objects provide important information for text-based image retrieval. As users are more likely to describe a scene from a real world perspective, usin…