2 papers
cs.LG2026
VP-VAE: Rethinking Vector Quantization via Adaptive Vector Perturbation
Linwei Zhai, Han Ding, Mingzhi Lin +5
Vector Quantized Variational Autoencoders (VQ-VAEs) are fundamental to modern generative modeling, yet they often suffer from training instability and "codebook collapse" due to th…
cs.SD2025
L3AC: Towards a Lightweight and Lossless Audio Codec
Linwei Zhai, Han Ding, Cui Zhao +4
Neural audio codecs have recently gained traction for their ability to compress high-fidelity audio and provide discrete tokens for generative modeling. However, leading approaches…