Showing eess.ASShow all
2 papers · 1 filter
eess.AS2026
Privacy-Preserving End-to-End Full-Duplex Speech Dialogue Models
Nikita Kuzmin, Tao Zhong, Jiajun Deng +6
End-to-end full-duplex speech models feed user audio through an always-on LLM backbone, yet the speaker privacy implications of their hidden representations remain unexamined. Foll…
eess.AS2025
CLEAR: Continuous Latent Autoregressive Modeling for High-quality and Low-latency Speech Synthesis
Chun Yat Wu, Jiajun Deng, Guinan Li +2
Autoregressive (AR) language models have emerged as powerful solutions for zero-shot text-to-speech (TTS) synthesis, capable of generating natural speech from a few seconds of audi…