Recent Advances in Medical Image Classification
arXiv:2506.04129 · doi:10.14569/ijacsa.2024.0150727
Abstract
Medical image classification is crucial for diagnosis and treatment, benefiting significantly from advancements in artificial intelligence. The paper reviews recent progress in the field, focusing on three levels of solutions: basic, specific, and applied. It highlights advances in traditional methods using deep learning models like Convolutional Neural Networks and Vision Transformers, as well as state-of-the-art approaches with Vision Language Models. These models tackle the issue of limited labeled data, and enhance and explain predictive results through Explainable Artificial Intelligence.
References in corpus (19)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- TransUNet: Transformers Make Strong Encoders for Medical Image Segmentation
- Predicting Cardiovascular Risk Factors from Retinal Fundus Photographs using Deep Learning
- The Medical Segmentation Decathlon
- MedViT: A Robust Vision Transformer for Generalized Medical Image Classification
- Contrastive Learning of Medical Visual Representations from Paired Images and Text
- XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models
- CosSIF: Cosine similarity-based image filtering to overcome low inter-class variation in synthetic medical image datasets
- LVM-Med: Learning Large-Scale Self-Supervised Vision Models for Medical Imaging via Second-order Graph Matching
- Image Projective Transformation Rectification with Synthetic Data for Smartphone-captured Chest X-ray Photos Classification
- MedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Texts
- A Comprehensive Study on Medical Image Segmentation using Deep Neural Networks
- MedKLIP: Medical Knowledge Enhanced Language-Image Pre-Training in Radiology
- DeViDe: Faceted medical knowledge for improved medical vision-language pre-training
- ECAMP: Entity-centered Context-aware Medical Vision Language Pre-training
- GazeGNN: A Gaze-Guided Graph Neural Network for Chest X-ray Classification
- MeDSLIP: Medical Dual-Stream Language-Image Pre-training with Pathology-Anatomy Semantic Alignment
- PM2: A New Prompting Multi-modal Model Paradigm for Few-shot Medical Image Classification