167 citations · 216 across the 15 of their papers we have counts for
25 papers
Point Cloud Recognition with Position-to-Structure Attention Transformers
Zheng Ding, James Hou, Zhuowen Tu
In this paper, we present Position-to-Structure Attention Transformers (PS-Former), a Transformer-based algorithm for 3D point cloud recognition. PS-Former deals with the challenge…
An In-depth Study of Stochastic Backpropagation
Jun Fang, Mingze Xu, Hao Chen +3
In this paper, we provide an in-depth study of Stochastic Backpropagation (SBP) when training deep neural networks for standard image classification and object detection tasks. Dur…
X-DETR: A Versatile Architecture for Instance-wise Vision-Language Tasks
Zhaowei Cai, Gukyeong Kwon, Avinash Ravichandran +4
In this paper, we study the challenging instance-wise vision-language tasks, where the free-form language is required to align with the objects instead of the whole image. To addre…
Text Spotting Transformers
Xiang Zhang, Yongwen Su, Subarna Tripathi +1
In this paper, we present TExt Spotting TRansformers (TESTR), a generic end-to-end text spotting framework using Transformers for text detection and recognition in the wild. TESTR…
MeMOT: Multi-Object Tracking with Memory
Jiarui Cai, Mingze Xu, Wei Li +4
We propose an online tracking algorithm that performs the object detection and data association under a common framework, capable of linking objects after a long time span. This is…
Contrastive Neighborhood Alignment
Pengkai Zhu, Zhaowei Cai, Yuanjun Xiong +4
We present Contrastive Neighborhood Alignment (CNA), a manifold learning approach to maintain the topology of learned features whereby data points that are mapped to nearby represe…