Publications (14)
Beyond Instance-Level Self-Supervision in 3D Multi-Modal Medical Imaging
Tan Pan, Shuhao Mei, Yixuan Sun +8
Self-supervised pre-training methods in medical imaging typically treat each individual as an isolated instance, learning representations through augmentation-based objectives or m…
Minimal Semantic Sufficiency Meets Unsupervised Domain Generalization
Tan Pan, Kaiyu Guo, Dongli Xu +8
The generalization ability of deep learning has been extensively studied in supervised settings, yet it remains less explored in unsupervised scenarios. Recently, the Unsupervised…
Aneumo: A Large-Scale Multimodal Aneurysm Dataset with Computational Fluid Dynamics Simulations and Deep Learning Benchmarks
Xigui Li, Yuanye Zhou, Feiyang Xiao +16
Intracranial aneurysms (IAs) are serious cerebrovascular lesions found in approximately 5\% of the general population. Their rupture may lead to high mortality. Current methods for…
Structure-aware Semantic Discrepancy and Consistency for 3D Medical Image Self-supervised Learning
Tan Pan, Zhaorui Tan, Kaiyu Guo +6
3D medical image self-supervised learning (mSSL) holds great promise for medical analysis. Effectively supporting broader applications requires considering anatomical structure var…
Boundary-aware Backward-Compatible Representation via Adversarial Learning in Image Retrieval
Tan Pan, Furong Xu, Xudong Yang +7
Image retrieval plays an important role in the Internet world. Usually, the core parts of mainstream visual retrieval systems include an online service of the embedding model and a…
Tracing the Heart's Pathways: ECG Representation Learning from a Cardiac Conduction Perspective
Tan Pan, Yixuan Sun, Chen Jiang +8
The multi-lead electrocardiogram (ECG) stands as a cornerstone of cardiac diagnosis. Recent strides in electrocardiogram self-supervised learning (eSSL) have brightened prospects f…
Towards a Universal 3D Medical Multi-modality Generalization via Learning Personalized Invariant Representation
Zhaorui Tan, Xi Yang, Tan Pan +8
Variations in medical imaging modalities and individual anatomical differences pose challenges to cross-modality generalization in multi-modal tasks. Existing methods often concent…
A Large-scale Comprehensive Dataset and Copy-overlap Aware Evaluation Protocol for Segment-level Video Copy Detection
Sifeng He, Xudong Yang, Chen Jiang +13
In this paper, we introduce VCSL (Video Copy Segment Localization), a new comprehensive segment-level annotated video copy dataset. Compared with existing copy detection datasets r…
Learning Segment Similarity and Alignment in Large-Scale Content Based Video Retrieval
Chen Jiang, Kaiming Huang, Sifeng He +9
With the explosive growth of web videos in recent years, large-scale Content-Based Video Retrieval (CBVR) becomes increasingly essential in video filtering, recommendation, and cop…
SD-MAD: Sign-Driven Few-shot Multi-Anomaly Detection in Medical Images
Kaiyu Guo, Tan Pan, Chen Jiang +5
Medical anomaly detection (AD) is crucial for early clinical intervention, yet it faces challenges due to limited access to high-quality medical imaging data, caused by privacy con…
Mind the Tool Failures: Achieving Synergistic Tool Gains for Medical Agents
Yunhui Gan, Tan Pan, Kaiyu Guo +5
Medical AI agents increasingly use external tools for diagnosis, treatment recommendation, and evidence retrieval, yet most existing approaches assume that task-appropriate tools a…
Improving Out-of-Distribution Detection via Dynamic Covariance Calibration
Kaiyu Guo, Zijian Wang, Tan Pan +2
Out-of-Distribution (OOD) detection is essential for the trustworthiness of AI systems. Methods using prior information (i.e., subspace-based methods) have shown effective performa…
Disco: Densely-overlapping Cell Instance Segmentation via Adjacency-aware Collaborative Coloring
Rui Sun, Yiwen Yang, Kaiyu Guo +8
Accurate cell instance segmentation is foundational for digital pathology analysis. Existing methods based on contour detection and distance mapping still face significant challeng…
Exploiting Layer Normalization Fine-tuning in Visual Transformer Foundation Models for Classification
Zhaorui Tan, Tan Pan, Kaizhu Huang +8
LayerNorm is pivotal in Vision Transformers (ViTs), yet its fine-tuning dynamics under data scarcity and domain shifts remain underexplored. This paper shows that shifts in LayerNo…