4 papers
Model Merging for Medical LVLMs: A Benchmark and a Winner-Take-All Approach
Lichao Mou, Shilan Zhang, Chunlei Li +7
Large vision-language models (LVLMs) can be adapted to specialized medical imaging tasks via parameter-efficient fine-tuning approaches such as low-rank adaptation (LoRA), leading…
Towards Real-World Ultrasound Understanding: Large Vision-Language Models from Multi-Image Examinations with Long-Form Reports
Bingcong Yan, Chunlei Li, Jingliang Hu +3
Large vision-language models (LVLMs) have achieved strong performance across many medical imaging tasks, yet their application to ultrasound remains limited due to its inherent com…
Camel: Frame-Level Bandwidth Estimation for Low-Latency Live Streaming under Video Bitrate Undershooting
Liming Liu, Zhidong Jia, Li Jiang +10
Low-latency live streaming (LLS) has emerged as a popular web application, with many platforms adopting real-time protocols such as WebRTC to minimize end-to-end latency. However,…
Ultrasound Image-to-Video Synthesis via Latent Dynamic Diffusion Models
Tingxiu Chen, Yilei Shi, Zixuan Zheng +4
Ultrasound video classification enables automated diagnosis and has emerged as an important research area. However, publicly available ultrasound video datasets remain scarce, hind…