2 citations · 4 across the 11 of their papers we have counts for
Showing 2025 · cs.CVShow all
3 papers · 2 filters
cs.CV2025
DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer
Xiaoya Tang, Bodong Zhang, Man Minh Ho +2
Despite the widespread adoption of transformers in medical applications, the exploration of multi-scale learning through transformers remains limited, while hierarchical representa…
cs.CV2025
A Comparison of Object Detection and Phrase Grounding Models in Chest X-ray Abnormality Localization using Eye-tracking Data
Elham Ghelichkhan, Tolga Tasdizen
Chest diseases rank among the most prevalent and dangerous global health issues. Object detection and phrase grounding deep learning models interpret complex radiology data to assi…
cs.CV2025
WeakSupCon: Weakly Supervised Contrastive Learning for Encoder Pre-training
Bodong Zhang, Hamid Manoochehri, Xiwen Li +2
Weakly supervised multiple instance learning (MIL) is a challenging task given that only bag-level labels are provided, while each bag typically contains multiple instances. This t…