665 citations · 802 across the 6 of their papers we have counts for
13 papers
Image Completion with Heterogeneously Filtered Spectral Hints
Xingqian Xu, Shant Navasardyan, Vahram Tadevosyan +3
Image completion with large-scale free-form missing regions is one of the most challenging tasks for the computer vision community. While researchers pursue better solutions, drawb…
Learning Sample Importance for Cross-Scenario Video Temporal Grounding
Peijun Bao, Yadong Mu
The task of temporal grounding aims to locate video moment in an untrimmed video, with a given sentence query. This paper for the first time investigates some superficial biases th…
Rethinking the Spatial Route Prior in Vision-and-Language Navigation
Xinzhe Zhou, Wei Liu, Yadong Mu
Vision-and-language navigation (VLN) is a trending topic which aims to navigate an intelligent agent to an expected position through natural language instructions. This work addres…
Poisoning MorphNet for Clean-Label Backdoor Attack to Point Clouds
Guiyu Tian, Wenhao Jiang, Wei Liu +1
This paper presents Poisoning MorphNet, the first backdoor attack method on point clouds. Conventional adversarial attack takes place in the inference stage, often fooling a model…
Informative Dropout for Robust Representation Learning: A Shape-bias Perspective
Baifeng Shi, Dinghuai Zhang, Qi Dai +3
Convolutional Neural Networks (CNNs) are known to rely more on local texture rather than global shape when making decisions. Recent work also indicates a close relationship between…
Weakly-Supervised Action Localization by Generative Attention Modeling
Baifeng Shi, Qi Dai, Yadong Mu +1
Weakly-supervised temporal action localization is a problem of learning an action localization model with only video-level action labeling available. The general framework largely…