2 papers
cs.CL2026
LLM-Bootstrapped Targeted Finding Guidance for Factual MLLM-based Medical Report Generation
Cunyuan Yang, Dejuan Song, Xiaotao Pang +7
The automatic generation of medical reports utilizing Multimodal Large Language Models (MLLMs) frequently encounters challenges related to factual instability, which may manifest a…
cs.CL2025
Multimodal Magic Elevating Depression Detection with a Fusion of Text and Audio Intelligence
Lindy Gan, Yifan Huang, Xiaoyang Gao +3
This study proposes an innovative multimodal fusion model based on a teacher-student architecture to enhance the accuracy of depression classification. Our designed model addresses…