INet: Inter-Intra-slice Interpolation Network for Medical Slice Synthesis
arXiv:2405.02857 · doi:10.1109/TMI.2024.3394033
Abstract
Medical imaging is limited by acquisition time and scanning equipment. CT and MR volumes, reconstructed with thicker slices, are anisotropic with high in-plane resolution and low through-plane resolution. We reveal an intriguing phenomenon that due to the mentioned nature of data, performing slice-wise interpolation from the axial view can yield greater benefits than performing super-resolution from other views. Based on this observation, we propose an Inter-Intra-slice Interpolation Network (INet), which fully explores information from high in-plane resolution and compensates for low through-plane resolution. The through-plane branch supplements the limited information contained in low through-plane resolution from high in-plane resolution and enables continual and diverse feature learning. In-plane branch transforms features to the frequency domain and enforces an equal learning opportunity for all frequency bands in a global context learning paradigm. We further propose a cross-view block to take advantage of the information from all three views online. Extensive experiments on two public datasets demonstrate the effectiveness of INet, and noticeably outperforms state-of-the-art super-resolution, video frame interpolation and slice interpolation methods by a large margin. We achieve 43.90dB in PSNR, with at least 1.14dB improvement under the upscale factor of 2 on MSD dataset with faster inference. Code is available at https://github.com/DeepMed-Lab-ECNU/Medical-Image-Reconstruction.
References in corpus (10)
- Adam: A Method for Stochastic Optimization
- Gaussian Error Linear Units (GELUs)
- The Medical Segmentation Decathlon
- CT Super-resolution GAN Constrained by the Identical, Residual, and Cycle Learning Ensemble(GAN-CIRCLE)
- The KiTS19 Challenge Data: 300 Kidney Tumor Cases with Clinical Context, CT Semantic Segmentations, and Surgical Outcomes
- Inter-slice Context Residual Learning for 3D Medical Image Segmentation
- An Arbitrary Scale Super-Resolution Approach for 3D MR Images via Implicit Neural Representation
- FDAN: Flow-guided Deformable Alignment Network for Video Super-Resolution
- RPLHR-CT Dataset and Transformer Baseline for Volumetric Super-Resolution from CT Scans
- Contrastive Learning for Local and Global Learning MRI Reconstruction