2 papers
cs.CV2025
Why Text Prevails: Vision May Undermine Multimodal Medical Decision Making
Siyuan Dai, Lunxiao Li, Kun Zhao +6
With the rapid progress of large language models (LLMs), advanced multimodal large language models (MLLMs) have demonstrated impressive zero-shot capabilities on vision-language ta…
cs.CV2025
Integrating Multi-scale and Multi-filtration Topological Features for Medical Image Classification
Pengfei Gu, Huimin Li, Haoteng Tang +6
Modern deep neural networks have shown remarkable performance in medical image classification. However, such networks either emphasize pixel-intensity features instead of fundament…