3 papers
eess.IV2026
Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation
Sunggu Kyung, Jinyoung Seo, Hyunseok Lim +7
Current CT report generation frameworks predominantly rely on global feature representations, often failing to capture region-specific details and potentially missing certain abnor…
cs.CV2025
MedErr-CT: A Visual Question Answering Benchmark for Identifying and Correcting Errors in CT Reports
Sunggu Kyung, Hyungbin Park, Jinyoung Seo +9
Computed Tomography (CT) plays a crucial role in clinical diagnosis, but the growing demand for CT examinations has raised concerns about diagnostic errors. While Multimodal Large…
cs.CL2025
MedOrchestra: A Hybrid Cloud-Local LLM Approach for Clinical Data Interpretation
Sihyeon Lee, Hyunjoo Song, Jong-chan Lee +6
Deploying large language models (LLMs) in clinical settings faces critical trade-offs: cloud LLMs, with their extensive parameters and superior performance, pose risks to sensitive…