4 citations · 6 across the 7 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2025★ 1 cited
Discrete Audio Tokens: More Than a Survey!
Pooneh Mousavi, Gallil Maimon, Adel Moumen +18
Discrete audio tokens are compact representations that aim to preserve perceptual quality, phonetic content, and speaker characteristics while enabling efficient storage and infere…
cs.SD2024★ 4 cited
Personalized Neural Speech Codec
Inseon Jang, Haici Yang, Wootaek Lim +2
In this paper, we propose a personalized neural speech codec, envisioning that personalization can reduce the model complexity or improve perceptual speech quality. Despite the com…
cs.SD2020
Non-local convolutional neural networks (nlcnn) for speaker recognition
Haici Yang, Hongda Mao, Ruirui Li +2
Speaker recognition is the process of identifying a speaker based on the voice. The technology has attracted more attention with the recent increase in popularity of smart voice as…