222 citations · 328 across the 139 of their papers we have counts for
1 paper · 2 filters
Georgios Ioannides, Adrian Kieback, Judah Goldfeder +5
Self-supervised speech encoders are predominantly trained by predicting discrete hard cluster IDs at masked positions, a recipe that collapses acoustic ambiguity at category bounda…