14 citations · 17 across the 18 of their papers we have counts for
19 papers
Physics-Aware Novel-View Acoustic Synthesis with Vision-Language Priors and 3D Acoustic Environment Modeling
Congyi Fan, Jian Guan, Youtian Lin +5
Spatial audio is essential for immersive experiences, yet novel-view acoustic synthesis (NVAS) remains challenging due to complex physical phenomena such as reflection, diffraction…
DualMark: Identifying Model and Training Data Origins in Generated Audio
Xuefeng Yang, Jian Guan, Feiyang Xiao +5
Existing watermarking methods for audio generative models only enable model-level attribution, allowing the identification of the originating generation model, but are unable to tr…
Attacking Voice Anonymization Systems with Augmented Feature and Speaker Identity Difference
Yanzhe Zhang, Zhonghao Bi, Feiyang Xiao +3
This study focuses on the First VoicePrivacy Attacker Challenge within the ICASSP 2025 Signal Processing Grand Challenge, which aims to develop speaker verification systems capable…
Disentangling Hierarchical Features for Anomalous Sound Detection Under Domain Shift
Jian Guan, Jiantong Tian, Qiaoxi Zhu +3
Anomalous sound detection (ASD) encounters difficulties with domain shift, where the sounds of machines in target domains differ significantly from those in source domains due to v…
Spectral-Temporal Fusion Representation for Person-in-Bed Detection
Xuefeng Yang, Shiheng Zhang, Jian Guan +3
This study is based on the ICASSP 2025 Signal Processing Grand Challenge's Accelerometer-Based Person-in-Bed Detection Challenge, which aims to determine bed occupancy using accele…
Graph-Enhanced Dual-Stream Feature Fusion with Pre-Trained Model for Acoustic Traffic Monitoring
Shitong Fan, Feiyang Xiao, Wenbo Wang +4
Microphone array techniques are widely used in sound source localization and smart city acoustic-based traffic monitoring, but these applications face significant challenges due to…