1 citations · 1 across the 4 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2025
StreamFlow: Streaming Flow Matching with Block-wise Guided Attention Mask for Speech Token Decoding
Dake Guo, Jixun Yao, Linhan Ma +2
Recent advancements in discrete token-based speech generation have highlighted the importance of token-to-waveform generation for audio quality, particularly in real-time interacti…
cs.SD2025
Boosting the Transferability of Audio Adversarial Examples with Acoustic Representation Optimization
Weifei Jin, Junjie Su, Hejia Wang +2
With the widespread application of automatic speech recognition (ASR) systems, their vulnerability to adversarial attacks has been extensively studied. However, most existing adver…
cs.SD2025★ 1 cited
The ICME 2025 Audio Encoder Capability Challenge
Junbo Zhang, Heinrich Dinkel, Qiong Song +8
This challenge aims to evaluate the capabilities of audio encoders, especially in the context of multi-task learning and real-world applications. Participants are invited to submit…