5 papers
RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity
Gaia A. Bertolino, Yuwei Zhang, Tong Xia +2
As conversational multimodal AI tools are increasingly adopted to process patient data for health assessment, robust benchmarks are needed to measure progress and expose failure mo…
RAMoEA-QA: Hierarchical Specialization for Robust Respiratory Audio Question Answering
Gaia A. Bertolino, Yuwei Zhang, Tong Xia +2
Conversational generative AI is increasingly explored in healthcare, where models must integrate heterogeneous patient signals and support diverse interaction styles while producin…
Wearable Foundation Models Should Go Beyond Static Encoders
Yu Yvonne Wu, Yuwei Zhang, Hyungjun Yoon +8
Wearable foundation models (WFMs), trained on large volumes of data collected by affordable, always-on devices, have demonstrated strong performance on short-term, well-defined hea…
Towards Open Respiratory Acoustic Foundation Models: Pretraining and Benchmarking
Yuwei Zhang, Tong Xia, Jing Han +6
Respiratory audio, such as coughing and breathing sounds, has predictive power for a wide range of healthcare applications, yet is currently under-explored. The main problem for th…
RespLLM: Unifying Audio and Text with Multimodal LLMs for Generalized Respiratory Health Prediction
Yuwei Zhang, Tong Xia, Aaqib Saeed +1
The high incidence and mortality rates associated with respiratory diseases underscores the importance of early screening. Machine learning models can automate clinical consultatio…