16 papers
CardioState-JEPA: Delay-Aware Cross-Modal Learning of a Shared Cardiac Representation
Hamza Shafiq, Hung Manh Pham, Bin Zhu +3
Electrocardiography (ECG), photoplethysmography (PPG), and phonocardiography (PCG) provide complementary views of the same cardiac cycle, yet existing cardiac foundation models are…
Unlocking In-Context Learning in Audio-Language Models from Decentralized Medical Audio
Ran Piao, Tsai-Ning Wang, Martijn den Dekker +4
Clinical audio diagnosis in low-resource settings requires models that identify conditions from minimal examples without large annotated corpora. We propose Federated Self-Contextu…
VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data
Di Zhu, Yu Yvonne Wu, Hong Jia +3
Wearable devices enable continuous monitoring of physiological signals such as ECG and PPG, but existing mHealth systems are largely limited to task-specific prediction pipelines o…
PulseLM: A Foundation Dataset and Benchmark for PPG-Text Learning
Hung Manh Pham, Jinyang Wu, Xiao Ma +6
Photoplethysmography (PPG) is a widely used non-invasive sensing modality for continuous cardiovascular and physiological monitoring across clinical, laboratory, and wearable setti…
Language Models as Semantic Teachers: Post-Training Alignment for Medical Audio Understanding
Tsai-Ning Wang, Lin-Lin Chen, Neil Zeghidour +1
Pre-trained audio models excel at detecting acoustic patterns in auscultation sounds but often fail to grasp their clinical significance, limiting their use and performance in diag…
Adaptive Test-Time Scaling for Zero-Shot Respiratory Audio Classification
Tsai-Ning Wang, Herman Teun den Dekker, Lin-Lin Chen +2
Automated respiratory audio analysis promises scalable, non-invasive disease screening, yet progress is limited by scarce labeled data and costly expert annotation. Zero-shot infer…