activity
20162025
most citedImproving Weakly Supervised Temporal Action Localization by Exploiting Multi-resolution Information in Temporal Domain

16 citations · 60 across the 13 of their papers we have counts for

collaborators
Showing cs.CVShow all

13 papers · 1 filter

cs.CV2025

Dense Video Captioning using Graph-based Sentence Summarization

Zhiwang Zhang, Dong Xu, Wanli Ouyang +1

Recently, dense video captioning has made attractive progress in detecting and captioning all events in a long untrimmed video. Despite promising results were achieved, most existi…

cs.CV2025

Progressive Modality Cooperation for Multi-Modality Domain Adaptation

Weichen Zhang, Dong Xu, Jing Zhang +1

In this work, we propose a new generic multi-modality domain adaptation framework called Progressive Modality Cooperation (PMC) to transfer the knowledge learned from the source do…

cs.CV2025

Self-Paced Collaborative and Adversarial Network for Unsupervised Domain Adaptation

Weichen Zhang, Dong Xu, Wanli Ouyang +1

This paper proposes a new unsupervised domain adaptation approach called Collaborative and Adversarial Network (CAN), which uses the domain-collaborative and domain-adversarial lea…

cs.CV202516 cited

Improving Weakly Supervised Temporal Action Localization by Exploiting Multi-resolution Information in Temporal Domain

Rui Su, Dong Xu, Luping Zhou +1

Weakly supervised temporal action localization is a challenging task as only the video-level annotation is available during the training process. To address this problem, we propos…

cs.CV20236 cited

Conflict-Based Cross-View Consistency for Semi-Supervised Semantic Segmentation

Zicheng Wang, Zhen Zhao, Xiaoxia Xing +3

Semi-supervised semantic segmentation (SSS) has recently gained increasing research interest as it can reduce the requirement for large-scale fully-annotated training data. The cur…

cs.CV2021

VoxelContext-Net: An Octree based Framework for Point Cloud Compression

Zizheng Que, Guo Lu, Dong Xu

In this paper, we propose a two-stage deep learning framework called VoxelContext-Net for both static and dynamic point cloud compression. Taking advantages of both octree based me…