self-supervised learning 2acoustic modeling 1language sensitivity 1low-resource languages 1multilingual 1multilingual speech 1neural audio codecs 1phone recognition 1speech representation 1
From the 2 of 10 linked papers with an AI index.
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2026
ESPnet3: Infrastructure for Scalable Speech and Audio Research in the Foundation Model Era
Masao Someki, Alexander Polok, Carlos Carvalho +14
Recent speech research involves increasingly large datasets, complex models, and diverse experimental workflows. However, existing frameworks require substantial engineering effort…
eess.AS2025
ESPnet-Codec: Comprehensive Training and Evaluation of Neural Codecs for Audio, Music, and Speech
Jiatong Shi, Jinchuan Tian, Yihan Wu +17
Neural codecs have become crucial to recent speech and audio generation research. In addition to signal compression capabilities, discrete codecs have also been found to enhance do…