activity
20182023
most citedDAOT: Domain-Agnostically Aligned Optimal Transport for Domain-Adaptive Crowd Counting

27 citations · 33 across the 4 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2023★ 27 cited

DAOT: Domain-Agnostically Aligned Optimal Transport for Domain-Adaptive Crowd Counting

Huilin Zhu, Jingling Yuan, Xian Zhong +3

Domain adaptation is commonly employed in crowd counting to bridge the domain gaps between different datasets. However, existing domain adaptation methods tend to focus on inter-da…

cs.CV2022★ 4 cited

Refined Semantic Enhancement towards Frequency Diffusion for Video Captioning

Xian Zhong, Zipeng Li, Shuqin Chen +3

Video captioning aims to generate natural language sentences that describe the given video accurately. Existing methods obtain favorable generation by exploring richer visual repre…

cs.CV2021

Visual-aware Attention Dual-stream Decoder for Video Captioning

Zhixin Sun, Xian Zhong, Shuqin Chen +2

Video captioning is a challenging task that captures different visual parts and describes them in sentences, for it requires visual and linguistic coherence. The attention mechanis…

cs.CV2020★ 2 cited

Complementing Representation Deficiency in Few-shot Image Classification: A Meta-Learning Approach

Xian Zhong, Cheng Gu, Wenxin Huang +3

Few-shot learning is a challenging problem that has attracted more and more attention recently since abundant training samples are difficult to obtain in practical applications. Me…

cs.CV2018

Image-to-Video Person Re-Identification by Reusing Cross-modal Embeddings

Zhongwei Xie, Lin Li, Xian Zhong +1

Image-to-video person re-identification identifies a target person by a probe image from quantities of pedestrian videos captured by non-overlapping cameras. Despite the great prog…