6 papers
Frequency Domain Image Translation: More Photo-realistic, Better Identity-preserving
Mu Cai, Hong Zhang, Huijuan Huang +3
Image-to-image translation has been revolutionized with GAN-based methods. However, existing methods lack the ability to preserve the identity of the source domain. As a result, sy…
Object-aware Feature Aggregation for Video Object Detection
Qichuan Geng, Hong Zhang, Na Jiang +3
We present an Object-aware Feature Aggregation (OFA) module for video object detection (VID). Our approach is motivated by the intriguing property that video-level object-aware kno…
Gated Path Selection Network for Semantic Segmentation
Qichuan Geng, Hong Zhang, Xiaojuan Qi +3
Semantic segmentation is a challenging task that needs to handle large scale variations, deformations and different viewpoints. In this paper, we develop a novel network named Gate…
Part-level Car Parsing and Reconstruction from Single Street View
Qichuan Geng, Hong Zhang, Xinyu Huang +5
Part information has been shown to be resistant to occlusions and viewpoint changes, which is beneficial for various vision-related tasks. However, we found very limited work in ca…
A Network Structure to Explicitly Reduce Confusion Errors in Semantic Segmentation
Qichuan Geng, Xinyu Huang, Zhong Zhou +1
Confusing classes that are ubiquitous in real world often degrade performance for many vision related applications like object detection, classification, and segmentation. The conf…
The ApolloScape Open Dataset for Autonomous Driving and its Application
Xinyu Huang, Peng Wang, Xinjing Cheng +3
Autonomous driving has attracted tremendous attention especially in the past few years. The key techniques for a self-driving car include solving tasks like 3D map construction, se…