most citedLeveraging Sound Source Trajectories for Universal Sound Separation

2 citations · 2 across the 4 of their papers we have counts for

collaborators

7 papers

eess.AS20262 cited

Leveraging Sound Source Trajectories for Universal Sound Separation

Donghang Wu, Xihong Wu, Tianshu Qu

Existing methods utilizing spatial information for sound source separation require prior knowledge of the direction of arrival (DOA) of the source or utilize estimated but imprecis…

cs.SD2026

SHB-AE: Spherical harmonic beamforming based Ambisonics encoding and upscaling method for smartphone microphone array

Yuhuan You, Yufan Qian, Tianshu Qu +2

With the rapid development of virtual reality (VR) and augmented reality (AR), spatial audio recording and reproduction have gained increasing research interest. Higher Order Ambis…

cs.SD2026

Flow-HOA: Generative Joint Optimization for Ambisonics Encoding via Flow Matching

Yuhuan You, Yufan Qian, Tianshu Qu +2

Higher-Order Ambisonics (HOA) encoding from sparse, irregular microphone arrays remains a critical challenge for consumer spatial audio capture in immersive communication and XR. W…

cs.SD2026

The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models

Yuhuan You, Lai Wei, Xihong Wu +1

Large audio-language models have made rapid progress in recognizing what is present in an audio clip, but spatial audio-language understanding still lacks a clear task interface. A…

eess.AS2025

Automotive sound field reproduction using deep optimization with spatial domain constraint

Yufan Qian, Tianshu Qu, Xihong Wu

Sound field reproduction with undistorted sound quality and precise spatial localization is desirable for automotive audio systems. However, the complexity of automotive cabin acou…

eess.AS2025

Cross-attention Inspired Selective State Space Models for Target Sound Extraction

Donghang Wu, Yiwen Wang, Xihong Wu +1

The Transformer model, particularly its cross-attention module, is widely used for feature fusion in target sound extraction which extracts the signal of interest based on given cl…