961 citations
- Nankai UniversityCN66 papers
- Chinese Academy of SciencesCN51 papers
- National University of SingaporeSG41 papers
- Tsinghua UniversityCN35 papers
- Peng Huanwu Center for Fundamental TheoryCN26 papers
- Peking UniversityCN23 papers
- Nanyang Technological UniversitySG22 papers
- Centre National de la Recherche ScientifiqueFR20 papers
- University of Science and Technology of ChinaCN19 papers
- University of Chinese Academy of SciencesCN18 papers
- Lanzhou UniversityCN17 papers
- City University of Hong KongHK16 papers
22 papers · 2 filters
Efficient Object-Level Visual Context Modeling for Multimodal Machine Translation: Masking Irrelevant Objects Helps Grounding
Dexin Wang, Deyi Xiong
Visual context provides grounding information for multimodal machine translation (MMT). However, previous MMT models and probing studies on visual features suggest that visual info…
PoNA: Pose-guided Non-local Attention for Human Pose Transfer
Kun Li, Jinsong Zhang, Yebin Liu +2
Human pose transfer, which aims at transferring the appearance of a given person to a target pose, is very challenging and important in many applications. Previous work ignores the…
Human Pose Transfer by Adaptive Hierarchical Deformation
Jinsong Zhang, Xingzi Liu, Kun Li
Human pose transfer, as a misaligned image generation task, is very challenging. Existing methods cannot effectively utilize the input information, which often fail to preserve the…
Transformer Guided Geometry Model for Flow-Based Unsupervised Visual Odometry
Xiangyu Li, Yonghong Hou, Pichao Wang +3
Existing unsupervised visual odometry (VO) methods either match pairwise images or integrate the temporal information using recurrent neural networks over a long sequence of images…
Fine-Grained Dynamic Head for Object Detection
Lin Song, Yanwei Li, Zhengkai Jiang +4
The Feature Pyramid Network (FPN) presents a remarkable approach to alleviate the scale variance in object representation by performing instance-level assignments. Nevertheless, th…
Rethinking Learnable Tree Filter for Generic Feature Transform
Lin Song, Yanwei Li, Zhengkai Jiang +5
The Learnable Tree Filter presents a remarkable approach to model structure-preserving relations for semantic segmentation. Nevertheless, the intrinsic geometric constraint forces…