papers

Publications (15)

cs.CV2022

R4D: Utilizing Reference Objects for Long-Range Distance Estimation

Yingwei Li, Tiffany Chen, Maya Kabkab +4

Estimating the distance of objects is a safety-critical task for autonomous driving. Focusing on short-range objects, existing methods and datasets neglect the equally important lo…

cs.CV2018

NISP: Pruning Networks using Neuron Importance Score Propagation

Ruichi Yu, Ang Li, Chun-Fu Chen +6

To reduce the significant redundancy in deep Convolutional Neural Networks (CNNs), most existing methods prune neurons by only considering statistics of an individual layer or two…

cs.CV2019

Layout-induced Video Representation for Recognizing Agent-in-Place Actions

Ruichi Yu, Hongcheng Wang, Ang Li +3

We address the recognition of agent-in-place actions, which are associated with agents who perform them and places where they occur, in the context of outdoor home surveillance. We…

cs.CV2017

Visual Relationship Detection with Internal and External Linguistic Knowledge Distillation

Ruichi Yu, Ang Li, Vlad I. Morariu +1

Understanding visual relationships involves identifying the subject, the object, and a predicate relating them. We leverage the strong correlations between the predicate and the (s…

cs.CV2019

Uncertainty Modeling of Contextual-Connections between Tracklets for Unconstrained Video-based Face Recognition

Jingxiao Zheng, Ruichi Yu, Jun-Cheng Chen +3

Unconstrained video-based face recognition is a challenging problem due to significant within-video variations caused by pose, occlusion and blur. To tackle this problem, an effect…

cs.CV2016

The Role of Context Selection in Object Detection

Ruichi Yu, Xi Chen, Vlad I. Morariu +1

We investigate the reasons why context in object detection has limited utility by isolating and evaluating the predictive power of different context cues under ideal conditions in…

cs.CV2018

ReMotENet: Efficient Relevant Motion Event Detection for Large-scale Home Surveillance Videos

Ruichi Yu, Hongcheng Wang, Larry S. Davis

This paper addresses the problem of detecting relevant motion caused by objects of interest (e.g., person and vehicles) in large scale home surveillance videos. The traditional met…

cs.CV2018

Modeling Local Geometric Structure of 3D Point Clouds using Geo-CNN

Shiyi Lan, Ruichi Yu, Gang Yu +1

Recent advances in deep convolutional neural networks (CNNs) have motivated researchers to adapt CNNs to directly model points in 3D point clouds. Modeling local structure has been…

cs.CV2018

Dynamic Zoom-in Network for Fast Object Detection in Large Images

Mingfei Gao, Ruichi Yu, Ang Li +2

We introduce a generic framework that reduces the computational cost of object detection while retaining accuracy for scenarios where objects with varied sizes appear in high resol…

cs.CV2020

SoDA: Multi-Object Tracking with Soft Data Association

Wei-Chih Hung, Henrik Kretzschmar, Tsung-Yi Lin +4

Robust multi-object tracking (MOT) is a prerequisite fora safe deployment of self-driving cars. Tracking objects, however, remains a highly challenging problem, especially in clutt…

cs.RO2024

STT: Stateful Tracking with Transformers for Autonomous Driving

Longlong Jing, Ruichi Yu, Xu Chen +19

Tracking objects in three-dimensional space is critical for autonomous driving. To ensure safety while driving, the tracker must be able to reliably track objects across frames and…

cs.CV2022

Depth Estimation Matters Most: Improving Per-Object Depth Estimation for Monocular 3D Detection and Tracking

Longlong Jing, Ruichi Yu, Henrik Kretzschmar +11

Monocular image-based 3D perception has become an active research area in recent years owing to its applications in autonomous driving. Approaches to monocular 3D perception includ…

cs.CV2018

C-WSL: Count-guided Weakly Supervised Localization

Mingfei Gao, Ang Li, Ruichi Yu +2

We introduce count-guided weakly supervised localization (C-WSL), an approach that uses per-class object count as a new form of supervision to improve weakly supervised localizatio…

cs.CV2018

VITON: An Image-based Virtual Try-on Network

Xintong Han, Zuxuan Wu, Zhe Wu +2

We present an image-based VIirtual Try-On Network (VITON) without using 3D information in any form, which seamlessly transfers a desired clothing item onto the corresponding region…

cs.CV2017

Generating Holistic 3D Scene Abstractions for Text-based Image Retrieval

Ang Li, Jin Sun, Joe Yue-Hei Ng +3

Spatial relationships between objects provide important information for text-based image retrieval. As users are more likely to describe a scene from a real world perspective, usin…