output
20062026
most citedDistance-IoU Loss: Faster and Better Learning for Bounding Box Regression

961 citations

Showing 2020 · cs.CVShow all

22 papers · 2 filters

cs.CV2020★ 1 cited

Efficient Object-Level Visual Context Modeling for Multimodal Machine Translation: Masking Irrelevant Objects Helps Grounding

Dexin Wang, Deyi Xiong

Visual context provides grounding information for multimodal machine translation (MMT). However, previous MMT models and probing studies on visual features suggest that visual info…

cs.CV2020★ 61 cited

PoNA: Pose-guided Non-local Attention for Human Pose Transfer

Kun Li, Jinsong Zhang, Yebin Liu +2

Human pose transfer, which aims at transferring the appearance of a given person to a target pose, is very challenging and important in many applications. Previous work ignores the…

cs.CV2020

Human Pose Transfer by Adaptive Hierarchical Deformation

Jinsong Zhang, Xingzi Liu, Kun Li

Human pose transfer, as a misaligned image generation task, is very challenging. Existing methods cannot effectively utilize the input information, which often fail to preserve the…

cs.CV2020

Transformer Guided Geometry Model for Flow-Based Unsupervised Visual Odometry

Xiangyu Li, Yonghong Hou, Pichao Wang +3

Existing unsupervised visual odometry (VO) methods either match pairwise images or integrate the temporal information using recurrent neural networks over a long sequence of images…

cs.CV2020★ 3 cited

Fine-Grained Dynamic Head for Object Detection

Lin Song, Yanwei Li, Zhengkai Jiang +4

The Feature Pyramid Network (FPN) presents a remarkable approach to alleviate the scale variance in object representation by performing instance-level assignments. Nevertheless, th…

cs.CV2020★ 6 cited

Rethinking Learnable Tree Filter for Generic Feature Transform

Lin Song, Yanwei Li, Zhengkai Jiang +5

The Learnable Tree Filter presents a remarkable approach to model structure-preserving relations for semantic segmentation. Nevertheless, the intrinsic geometric constraint forces…