2 citations · 2 across the 2 of their papers we have counts for
5 papers
Leveraging Sound Source Trajectories for Universal Sound Separation
Donghang Wu, Xihong Wu, Tianshu Qu
Existing methods utilizing spatial information for sound source separation require prior knowledge of the direction of arrival (DOA) of the source or utilize estimated but imprecis…
The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models
Yuhuan You, Lai Wei, Xihong Wu +1
Large audio-language models have made rapid progress in recognizing what is present in an audio clip, but spatial audio-language understanding still lacks a clear task interface. A…
Automotive sound field reproduction using deep optimization with spatial domain constraint
Yufan Qian, Tianshu Qu, Xihong Wu
Sound field reproduction with undistorted sound quality and precise spatial localization is desirable for automotive audio systems. However, the complexity of automotive cabin acou…
Cross-attention Inspired Selective State Space Models for Target Sound Extraction
Donghang Wu, Yiwen Wang, Xihong Wu +1
The Transformer model, particularly its cross-attention module, is widely used for feature fusion in target sound extraction which extracts the signal of interest based on given cl…
DENSE: Dynamic Embedding Causal Target Speech Extraction
Yiwen Wang, Zeyu Yuan, Xihong Wu
Target speech extraction (TSE) focuses on extracting the speech of a specific target speaker from a mixture of signals. Existing TSE models typically utilize static embeddings as c…