6 citations · 28 across the 20 of their papers we have counts for
12 papers
Learning Partial Correlation based Deep Visual Representation for Image Classification
Saimunur Rahman, Piotr Koniusz, Lei Wang +3
Visual representation based on covariance matrix has demonstrates its efficacy for image classification by characterising the pairwise correlation of different channels in convolut…
METransformer: Radiology Report Generation by Transformer with Multiple Learnable Expert Tokens
Zhanyu Wang, Lingqiao Liu, Lei Wang +1
In clinical scenarios, multi-specialist consultation could significantly benefit the diagnosis, especially for intricate cases. This inspires us to explore a "multi-expert joint di…
Q2ATransformer: Improving Medical VQA via an Answer Querying Decoder
Yunyi Liu, Zhanyu Wang, Dong Xu +1
Medical Visual Question Answering (VQA) systems play a supporting role to understand clinic-relevant information carried by medical images. The questions to a medical image include…
Task-Oriented Multi-Modal Mutual Leaning for Vision-Language Models
Sifan Long, Zhen Zhao, Junkun Yuan +5
Prompt learning has become one of the most efficient paradigms for adapting large pre-trained vision-language models to downstream tasks. Current state-of-the-art methods, like CoO…
Conflict-Based Cross-View Consistency for Semi-Supervised Semantic Segmentation
Zicheng Wang, Zhen Zhao, Xiaoxia Xing +3
Semi-supervised semantic segmentation (SSS) has recently gained increasing research interest as it can reduce the requirement for large-scale fully-annotated training data. The cur…
Neural Vector Fields: Implicit Representation by Explicit Learning
Xianghui Yang, Guosheng Lin, Zhenghao Chen +1
Deep neural networks (DNNs) are widely applied for nowadays 3D surface reconstruction tasks and such methods can be further divided into two categories, which respectively warp tem…