187 citations · 380 across the 11 of their papers we have counts for
11 papers
Masked Vision-Language Transformers for Scene Text Recognition
Jie Wu, Ying Peng, Shengming Zhang +2
Scene text recognition (STR) enables computers to recognize and read the text in various real-world scenes. Recent STR models benefit from taking linguistic information in addition…
Dyn-Backdoor: Backdoor Attack on Dynamic Link Prediction
Jinyin Chen, Haiyang Xiong, Haibin Zheng +3
Dynamic link prediction (DLP) makes graph prediction based on historical information. Since most DLP methods are highly dependent on the training data to achieve satisfying predict…
Learning Disentangled Representation Implicitly via Transformer for Occluded Person Re-Identification
Mengxi Jia, Xinhua Cheng, Shijian Lu +1
Person re-identification (re-ID) under various occlusions has been a long-standing challenge as person images with different types of occlusions often suffer from misalignment in i…
Super-Resolving Compressed Video in Coding Chain
Dewang Hou, Yang Zhao, Yuyao Ye +3
Scaling and lossy coding are widely used in video transmission and storage. Previous methods for enhancing the resolution of such videos often ignore the inherent interference betw…
A comprehensive survey on point cloud registration
Xiaoshui Huang, Guofeng Mei, Jian Zhang +1
Registration is a transformation estimation problem between two point clouds, which has a unique and critical role in numerous computer vision applications. The developments of opt…
Multiple Instance Segmentation in Brachial Plexus Ultrasound Image Using BPMSegNet
Yi Ding, Qiqi Yang, Guozheng Wu +2
The identification of nerve is difficult as structures of nerves are challenging to image and to detect in ultrasound images. Nevertheless, the nerve identification in ultrasound i…