3 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 3 cited
Language Adaptive Weight Generation for Multi-task Visual Grounding
Wei Su, Peihan Miao, Huanzhang Dou +4
Although the impressive performance in visual grounding, the prevailing approaches usually exploit the visual backbone in a passive way, i.e., the visual backbone extracts features…
cs.CV2022★ 1 cited
Unified Normalization for Accelerating and Stabilizing Transformers
Qiming Yang, Kai Zhang, Chaoxiang Lan +5
Solid results from Transformers have made them prevailing architectures in various natural language and vision tasks. As a default component in Transformers, Layer Normalization (L…