2 papers
eess.IV2025
Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts
Peixuan Ge, Tongkun Su, Faqin Lv +8
Ultrasound (US) report generation is a challenging task due to the variability of US images, operator dependence, and the need for standardized text. Unlike X-ray and CT, US imagin…
cs.CV2024
Design as Desired: Utilizing Visual Question Answering for Multimodal Pre-training
Tongkun Su, Jun Li, Xi Zhang +6
Multimodal pre-training demonstrates its potential in the medical domain, which learns medical visual representations from paired medical reports. However, many pre-training tasks…