8 citations · 35 across the 37 of their papers we have counts for
14 papers · 1 filter
Multimodal Variational Autoencoder: a Barycentric View
Peijie Qiu, Wenhui Zhu, Sayantan Kumar +6
Multiple signal modalities, such as vision and sounds, are naturally present in real-world phenomena. Recently, there has been growing interest in learning generative models, in pa…
Geographical Information Alignment Boosts Traffic Analysis via Transpose Cross-attention
Xiangyu Jiang, Xiwen Chen, Hao Wang +1
Traffic accident prediction is crucial for enhancing road safety and mitigating congestion, and recent Graph Neural Networks (GNNs) have shown promise in modeling the inherent grap…
Many-MobileNet: Multi-Model Augmentation for Robust Retinal Disease Classification
Hao Wang, Wenhui Zhu, Xuanzhao Dong +9
In this work, we propose Many-MobileNet, an efficient model fusion strategy for retinal disease classification using lightweight CNN architecture. Our method addresses key challeng…
RobustFormer: Noise-Robust Pre-training for images and videos
Ashish Bastola, Nishant Luitel, Hao Wang +3
While deep learning-based models like transformers, have revolutionized time-series and vision tasks, they remain highly susceptible to noise and often overfit on noisy patterns ra…
RBAD: A Dataset and Benchmark for Retinal Vessels Branching Angle Detection
Hao Wang, Wenhui Zhu, Jiayou Qin +5
Detecting retinal image analysis, particularly the geometrical features of branching points, plays an essential role in diagnosing eye diseases. However, existing methods used for…
DGR-MIL: Exploring Diverse Global Representation in Multiple Instance Learning for Whole Slide Image Classification
Wenhui Zhu, Xiwen Chen, Peijie Qiu +3
Multiple instance learning (MIL) stands as a powerful approach in weakly supervised learning, regularly employed in histological whole slide image (WSI) classification for detectin…