12 citations · 18 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 4 cited
WITT: A Wireless Image Transmission Transformer for Semantic Communications
Ke Yang, Sixian Wang, Jincheng Dai +3
In this paper, we aim to redesign the vision Transformer (ViT) as a new backbone to realize semantic image transmission, termed wireless image transmission transformer (WITT). Prev…
cs.CL2022★ 2 cited
Attract me to Buy: Advertisement Copywriting Generation with Multimodal Multi-structured Information
Zhipeng Zhang, Xinglin Hou, Kai Niu +5
Recently, online shopping has gradually become a common way of shopping for people all over the world. Wonderful merchandise advertisements often attract more people to buy. These…
cs.CV2019★ 12 cited
Improving Description-based Person Re-identification by Multi-granularity Image-text Alignments
Kai Niu, Yan Huang, Wanli Ouyang +1
Description-based person re-identification (Re-id) is an important task in video surveillance that requires discriminative cross-modal representations to distinguish different peop…