19 citations · 20 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 19 cited
LT-ViT: A Vision Transformer for multi-label Chest X-ray classification
Umar Marikkar, Sara Atito, Muhammad Awais +1
Vision Transformers (ViTs) are widely adopted in medical imaging tasks, and some existing efforts have been directed towards vision-language training for Chest X-rays (CXRs). Howev…
cs.CV2023★ 1 cited
SCD-Net: Spatiotemporal Clues Disentanglement Network for Self-supervised Skeleton-based Action Recognition
Cong Wu, Xiao-Jun Wu, Josef Kittler +4
Contrastive learning has achieved great success in skeleton-based action recognition. However, most existing approaches encode the skeleton sequences as entangled spatiotemporal re…
cs.CV2023
Masked Momentum Contrastive Learning for Zero-shot Semantic Understanding
Jiantao Wu, Shentong Mo, Muhammad Awais +3
Self-supervised pretraining (SSP) has emerged as a popular technique in machine learning, enabling the extraction of meaningful feature representations without labelled data. In th…