Showing eess.ASShow all
2 papers · 1 filter
eess.AS2026
VoxWatermark: A Large-Scale Benchmark for Audio Watermark Detection under Perturbations
Farnaz Sedaghati, Yuxi Wang, Zicheng Weng +1
With the rapid deployment of speech generation systems in open environments, providing verifiable source attribution and copyright accountability for audio content has become criti…
eess.AS2026
Spatial-Omni: Spatial Audio Understanding Integration in Multimodal LLMs via FOA Encoding
Zhiyuan Zhu, Yixuan Chen, Yiwen Shao +13
Recent multimodal large language models mainly process audio as monaural signals, thereby discarding the spatial cues contained in spatial audio for sound localization, spatial rel…