activity
20152022
most citedTraining Deeper Convolutional Networks with Deep Supervision

167 citations · 216 across the 15 of their papers we have counts for

collaborators

25 papers

cs.CV2022

Point Cloud Recognition with Position-to-Structure Attention Transformers

Zheng Ding, James Hou, Zhuowen Tu

In this paper, we present Position-to-Structure Attention Transformers (PS-Former), a Transformer-based algorithm for 3D point cloud recognition. PS-Former deals with the challenge…

cs.CV20221 cited

An In-depth Study of Stochastic Backpropagation

Jun Fang, Mingze Xu, Hao Chen +3

In this paper, we provide an in-depth study of Stochastic Backpropagation (SBP) when training deep neural networks for standard image classification and object detection tasks. Dur…

cs.CV20223 cited

X-DETR: A Versatile Architecture for Instance-wise Vision-Language Tasks

Zhaowei Cai, Gukyeong Kwon, Avinash Ravichandran +4

In this paper, we study the challenging instance-wise vision-language tasks, where the free-form language is required to align with the objects instead of the whole image. To addre…

cs.CV2022

Text Spotting Transformers

Xiang Zhang, Yongwen Su, Subarna Tripathi +1

In this paper, we present TExt Spotting TRansformers (TESTR), a generic end-to-end text spotting framework using Transformers for text detection and recognition in the wild. TESTR…

cs.CV2022

MeMOT: Multi-Object Tracking with Memory

Jiarui Cai, Mingze Xu, Wei Li +4

We propose an online tracking algorithm that performs the object detection and data association under a common framework, capable of linking objects after a long time span. This is…

cs.LG20223 cited

Contrastive Neighborhood Alignment

Pengkai Zhu, Zhaowei Cai, Yuanjun Xiong +4

We present Contrastive Neighborhood Alignment (CNA), a manifold learning approach to maintain the topology of learned features whereby data points that are mapped to nearby represe…