3 papers
eess.AS2026
An Ultra-Low Latency, End-to-End Streaming Speech Synthesis Architecture via Block-Wise Generation and Depth-Wise Codec Decoding
Tianhui Su, Tien-Ping Tan, Salima Mdhaffar +2
Real-time speech synthesis requires balancing inference latency and acoustic fidelity for interactive applications. Conventional continuous text-to-speech pipelines require computa…
cs.SD2025
Sparse deepfake detection promotes better disentanglement
Antoine Teissier, Marie Tahon, Nicolas Dugué +1
Due to the rapid progress of speech synthesis, deepfake detection has become a major concern in the speech processing community. Because it is a critical task, systems must not onl…
cs.SD2025
Multi-level SSL Feature Gating for Audio Deepfake Detection
Hoan My Tran, Damien Lolive, Aghilas Sini +3
Recent advancements in generative AI, particularly in speech synthesis, have enabled the generation of highly natural-sounding synthetic speech that closely mimics human voices. Wh…