310 citations · 310 across the 2 of their papers we have counts for
1 paper · 1 filter
Alexei Baevski, Steffen Schneider, Michael Auli
We propose vq-wav2vec to learn discrete representations of audio segments through a wav2vec-style self-supervised context prediction task. The algorithm uses either a gumbel softma…