activity
20162024
most citedTowards Lightweight Transformer via Group-wise Transformation for Vision-and-Language Tasks

66 citations · 161 across the 21 of their papers we have counts for

collaborators
Showing cs.CVShow all

33 papers · 1 filter

cs.CV2024

Image Captioning via Dynamic Path Customization

Yiwei Ma, Jiayi Ji, Xiaoshuai Sun +4

This paper explores a novel dynamic network for vision and language tasks, where the inferring structure is customized on the fly for different inputs. Most previous state-of-the-a…

cs.CV2023★ 2 cited

Latent Feature Relation Consistency for Adversarial Robustness

Xingbin Liu, Huafeng Kuang, Hong Liu +3

Deep neural networks have been applied in many computer vision tasks and achieved state-of-the-art performance. However, misclassification will occur when DNN predicts adversarial…

cs.CV2023★ 1 cited

CAT:Collaborative Adversarial Training

Xingbin Liu, Huafeng Kuang, Xianming Lin +2

Adversarial training can improve the robustness of neural networks. Previous methods focus on a single adversarial training strategy and do not consider the model property trained…

cs.CV2023★ 3 cited

Attention Disturbance and Dual-Path Constraint Network for Occluded Person Re-identification

Jiaer Xia, Lei Tan, Pingyang Dai +3

Occluded person re-identification (Re-ID) aims to address the potential occlusion problem when matching occluded or holistic pedestrians from different camera views. Many methods u…

cs.CV2023

Spectral Aware Softmax for Visible-Infrared Person Re-Identification

Lei Tan, Pingyang Dai, Qixiang Ye +3

Visible-infrared person re-identification (VI-ReID) aims to match specific pedestrian images from different modalities. Although suffering an extra modality discrepancy, existing m…

cs.CV2023★ 10 cited

Exploring Invariant Representation for Visible-Infrared Person Re-Identification

Lei Tan, Yukang Zhang, Shengmei Shen +5

Cross-spectral person re-identification, which aims to associate identities to pedestrians across different spectra, faces a main challenge of the modality discrepancy. In this pap…