27 citations · 33 across the 4 of their papers we have counts for
5 papers · 1 filter
DAOT: Domain-Agnostically Aligned Optimal Transport for Domain-Adaptive Crowd Counting
Huilin Zhu, Jingling Yuan, Xian Zhong +3
Domain adaptation is commonly employed in crowd counting to bridge the domain gaps between different datasets. However, existing domain adaptation methods tend to focus on inter-da…
Refined Semantic Enhancement towards Frequency Diffusion for Video Captioning
Xian Zhong, Zipeng Li, Shuqin Chen +3
Video captioning aims to generate natural language sentences that describe the given video accurately. Existing methods obtain favorable generation by exploring richer visual repre…
Visual-aware Attention Dual-stream Decoder for Video Captioning
Zhixin Sun, Xian Zhong, Shuqin Chen +2
Video captioning is a challenging task that captures different visual parts and describes them in sentences, for it requires visual and linguistic coherence. The attention mechanis…
Complementing Representation Deficiency in Few-shot Image Classification: A Meta-Learning Approach
Xian Zhong, Cheng Gu, Wenxin Huang +3
Few-shot learning is a challenging problem that has attracted more and more attention recently since abundant training samples are difficult to obtain in practical applications. Me…
Image-to-Video Person Re-Identification by Reusing Cross-modal Embeddings
Zhongwei Xie, Lin Li, Xian Zhong +1
Image-to-video person re-identification identifies a target person by a probe image from quantities of pedestrian videos captured by non-overlapping cameras. Despite the great prog…