12 citations · 32 across the 6 of their papers we have counts for
13 papers
3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language Models
Hao Chen, Wei Zhao, Yingli Li +8
Medical image analysis is crucial in modern radiological diagnostics, especially given the exponential growth in medical imaging data. The demand for automated report generation sy…
HARP: Human-Assisted Regrouping with Permutation Invariant Critic for Multi-Agent Reinforcement Learning
Huawen Hu, Enze Shi, Chenxi Yue +7
Human-in-the-loop reinforcement learning integrates human expertise to accelerate agent learning and provide critical guidance and feedback in complex fields. However, many existin…
Identifying Influential nodes in Brain Networks via Self-Supervised Graph-Transformer
Yanqing Kang, Di Zhu, Haiyang Zhang +9
Studying influential nodes (I-nodes) in brain networks is of great significance in the field of brain imaging. Most existing studies consider brain connectivity hubs as I-nodes. Ho…
ModalityMirror: Improving Audio Classification in Modality Heterogeneity Federated Learning with Multimodal Distillation
Tiantian Feng, Tuo Zhang, Salman Avestimehr +1
Multimodal Federated Learning frequently encounters challenges of client modality heterogeneity, leading to undesired performances for secondary modality in multimodal learning. It…
A Comprehensive Review of Multimodal Large Language Models: Performance and Challenges Across Different Tasks
Jiaqi Wang, Hanqi Jiang, Yiheng Liu +21
In an era defined by the explosive growth of data and rapid technological advancements, Multimodal Large Language Models (MLLMs) stand at the forefront of artificial intelligence (…
Potential of Multimodal Large Language Models for Data Mining of Medical Images and Free-text Reports
Yutong Zhang, Yi Pan, Tianyang Zhong +11
Medical images and radiology reports are crucial for diagnosing medical conditions, highlighting the importance of quantitative analysis for clinical decision-making. However, the…