2 papers
cs.SD2025
Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations
Yichen Han, Xiaoyang Hao, Keming Chen +25
Text-to-speech (TTS) synthesis has seen renewed progress under the discrete modeling paradigm. Existing autoregressive approaches often rely on single-codebook representations, whi…
cs.CV2024
A Practical Gated Recurrent Transformer Network Incorporating Multiple Fusions for Video Denoising
Kai Guo, Seungwon Choi, Jongseong Choi +1
State-of-the-art (SOTA) video denoising methods employ multi-frame simultaneous denoising mechanisms, resulting in significant delays (e.g., 16 frames), making them impractical for…