papers

Publications (43)

eess.IV2024

Prompted Contextual Transformer for Incomplete-View CT Reconstruction

Chenglong Ma, Zilong Li, Junjun He +3

Incomplete-view computed tomography (CT) can shorten the data acquisition time and allow scanning of large objects, including sparse-view and limited-angle scenarios, each with var…

cs.CV2025

GMAI-VL-R1: Harnessing Reinforcement Learning for Multimodal Medical Reasoning

Yanzhou Su, Tianbin Li, Jiyao Liu +15

Recent advances in general medical AI have made significant strides, but existing models often lack the reasoning capabilities needed for complex medical decision-making. This pape…

cs.CV2026

Bridging Brain and Semantics: A Hierarchical Framework for Semantically Enhanced fMRI-to-Video Reconstruction

Yujie Wei, Chenglong Ma, Jianxiong Gao +6

Reconstructing dynamic visual experiences as videos from functional magnetic resonance imaging (fMRI) is pivotal for advancing the understanding of neural processes. However, curre…

eess.IV2024

SegBook: A Simple Baseline and Cookbook for Volumetric Medical Image Segmentation

Jin Ye, Ying Chen, Yanjun Li +7

Computed Tomography (CT) is one of the most popular modalities for medical imaging. By far, CT images have contributed to the largest publicly available datasets for volumetric med…

eess.IV2023

FreeSeed: Frequency-band-aware and Self-guided Network for Sparse-view CT Reconstruction

Chenglong Ma, Zilong Li, Junping Zhang +2

Sparse-view computed tomography (CT) is a promising solution for expediting the scanning process and mitigating radiation exposure to patients, the reconstructed images, however, c…

cs.CV2025

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Tianbin Li, Yanzhou Su, Wei Li +15

Despite significant advancements in general AI, its effectiveness in the medical domain is limited by the lack of specialized medical knowledge. To address this, we formulate GMAI-…