2 papers
cs.CV2026
Seizure-Semiology-Suite (S3): A Clinically Multimodal Dataset, Benchmark, and Models for Seizure Semiology Understanding
Lina Zhang, Tonmoy Monsoor, Peizheng Li +23
While Multimodal Large Language Models (MLLMs) have demonstrated remarkable proficiency in general video understanding, their capacity to interpret involuntary, and spatio-temporal…
eess.IV2025
MSV-Mamba: A Multiscale Vision Mamba Network for Echocardiography Segmentation
Xiaoxian Yang, Qi Wang, Kaiqi Zhang +3
Ultrasound imaging frequently encounters challenges, such as those related to elevated noise levels, diminished spatiotemporal resolution, and the complexity of anatomical structur…