3 citations · 3 across the 6 of their papers we have counts for
4 papers · 1 filter
SAFIRE: Safety-Critical Benchmark for Fine-grained Fire and Smoke Understanding in Multimodal LLMs
Pengfei Li, Naufal Suryanto, Sicheng Zhang +2
Multimodal Large Language Models (MLLMs) show strong progress on vision-language tasks, yet their reliability in safety-critical settings remains underexplored. Fire-smoke understa…
Experience-Guided Self-Adaptive Cascaded Agents for Breast Cancer Screening and Diagnosis with Reduced Biopsy Referrals
Pramit Saha, Mohammad Alsharid, Joshua Strong +1
We propose an experience-guided cascaded multi-agent framework for Breast Ultrasound Screening and Diagnosis, called BUSD-Agent, that aims to reduce diagnostic escalation and unnec…
Show from Tell: Audio-Visual Modelling in Clinical Settings
Jianbo Jiao, Mohammad Alsharid, Lior Drukker +3
Auditory and visual signals usually present together and correlate with each other, not only in natural environments but also in clinical settings. However, the audio-visual modell…
Self-supervised Contrastive Video-Speech Representation Learning for Ultrasound
Jianbo Jiao, Yifan Cai, Mohammad Alsharid +3
In medical imaging, manual annotations can be expensive to acquire and sometimes infeasible to access, making conventional deep learning-based models difficult to scale. As a resul…