1 paper
Khanh Le, Kiet Anh Hoang, Bao Nguyen +5
We present ViP-VL, an efficient Vietnamese Self-supervised speech Pretraining model leveraging Vector-quantization Learning. To bridge the gap between high-resolution audio and eff…