9 citations · 12 across the 5 of their papers we have counts for
5 papers
Diverse Instance Discovery: Vision-Transformer for Instance-Aware Multi-Label Image Recognition
Yunqing Hu, Xuan Jin, Yin Zhang +5
Previous works on multi-label image recognition (MLIR) usually use CNNs as a starting point for research. In this paper, we take pure Vision Transformer (ViT) as the research base…
DRDF: Determining the Importance of Different Multimodal Information with Dual-Router Dynamic Framework
Haiwen Hong, Xuan Jin, Yin Zhang +4
In multimodal tasks, we find that the importance of text and image modal information is different for different input cases, and for this motivation, we propose a high-performance…
RAMS-Trans: Recurrent Attention Multi-scale Transformer forFine-grained Image Recognition
Yunqing Hu, Xuan Jin, Yin Zhang +4
In fine-grained image recognition (FGIR), the localization and amplification of region attention is an important factor, which has been explored a lot by convolutional neural netwo…
The Open Brands Dataset: Unified brand detection and recognition at scale
Xuan Jin, Wei Su, Rong Zhang +2
Intellectual property protection(IPP) have received more and more attention recently due to the development of the global e-commerce platforms. brand recognition plays a significan…
Heuristic Domain Adaptation
Shuhao Cui, Xuan Jin, Shuhui Wang +2
In visual domain adaptation (DA), separating the domain-specific characteristics from the domain-invariant representations is an ill-posed problem. Existing methods apply different…