1 paper
Chihiro Taguchi, Ãric Le Ferrand, Hirosi Nakagawa +4
Modern pretrained self-supervised automatic speech recognition models are trained on large-scale audio data to encode speech into contextualized representations. However, their tra…