3 papers
cs.LG2026
OpenAVS: Training-Free Open-Vocabulary Audio Visual Segmentation with Foundational Models
Shengkai Chen, Yifang Yin, Jinming Cao +3
Audio-visual segmentation aims to separate sounding objects from videos by predicting pixel-level masks based on audio signals. Existing methods primarily concentrate on closed-set…
cs.CV2025
SimCast: Enhancing Precipitation Nowcasting with Short-to-Long Term Knowledge Distillation
Yifang Yin, Shengkai Chen, Yiyao Li +4
Precipitation nowcasting predicts future radar sequences based on current observations, which is a highly challenging task driven by the inherent complexity of the Earth system. Ac…
cs.LG2024
Prompt-Based Spatio-Temporal Graph Transfer Learning
Junfeng Hu, Xu Liu, Zhencheng Fan +4
Spatio-temporal graph neural networks have proven efficacy in capturing complex dependencies for urban computing tasks such as forecasting and kriging. Yet, their performance is co…