1 paper
Chunhui Zhang, Li Liu, Jialin Gao +5
Transformer has recently demonstrated great potential in improving vision-language (VL) tracking algorithms. However, most of the existing VL trackers rely on carefully designed me…