5 papers
Dense Video Captioning using Graph-based Sentence Summarization
Zhiwang Zhang, Dong Xu, Wanli Ouyang +1
Recently, dense video captioning has made attractive progress in detecting and captioning all events in a long untrimmed video. Despite promising results were achieved, most existi…
Progressive Modality Cooperation for Multi-Modality Domain Adaptation
Weichen Zhang, Dong Xu, Jing Zhang +1
In this work, we propose a new generic multi-modality domain adaptation framework called Progressive Modality Cooperation (PMC) to transfer the knowledge learned from the source do…
Self-Paced Collaborative and Adversarial Network for Unsupervised Domain Adaptation
Weichen Zhang, Dong Xu, Wanli Ouyang +1
This paper proposes a new unsupervised domain adaptation approach called Collaborative and Adversarial Network (CAN), which uses the domain-collaborative and domain-adversarial lea…
Progressive Cross-Stream Cooperation in Spatial and Temporal Domain for Action Localization
Rui Su, Dong Xu, Luping Zhou +1
Spatio-temporal action localization consists of three levels of tasks: spatial localization, action classification, and temporal localization. In this work, we propose a new progre…
Improving Weakly Supervised Temporal Action Localization by Exploiting Multi-resolution Information in Temporal Domain
Rui Su, Dong Xu, Luping Zhou +1
Weakly supervised temporal action localization is a challenging task as only the video-level annotation is available during the training process. To address this problem, we propos…