3 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 3 cited
Uni-Perceiver v2: A Generalist Model for Large-Scale Vision and Vision-Language Tasks
Hao Li, Jinguo Zhu, Xiaohu Jiang +8
Despite the remarkable success of foundation models, their task-specific fine-tuning paradigm makes them inconsistent with the goal of general perception modeling. The key to elimi…
cs.CV2021★ 1 cited
Guiding Query Position and Performing Similar Attention for Transformer-Based Detection Heads
Xiaohu Jiang, Ze Chen, Zhicheng Wang +2
After DETR was proposed, this novel transformer-based detection paradigm which performs several cross-attentions between object queries and feature maps for predictions has subsequ…