language sensitivity 1multilingual 1neural audio codecs 1self-supervised learning 1speech representation 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.SD2026
Dissecting Sensitivity to Training Language in Self-Supervised Speech Learning Using Neural Audio Codec Tokens
Daigo Takizawa, Tomohiko Nakamura, Samuele Cornell +3
The paper investigates how the language used to train neural audio codecs and self‑supervised speech models affects performance, finding that codec training language has little imp…
eess.AS2025
IdolSongsJp Corpus: A Multi-Singer Song Corpus in the Style of Japanese Idol Groups
Hitoshi Suda, Junya Koguchi, Shunsuke Yoshida +3
Japanese idol groups, comprising performers known as "idols," are an indispensable part of Japanese pop culture. They frequently appear in live concerts and television programs, en…
eess.AS2025
Discrete Speech Unit Extraction via Independent Component Analysis
Tomohiko Nakamura, Kwanghee Choi, Keigo Hojo +3
Self-supervised speech models (S3Ms) have become a common tool for the speech processing community, leveraging representations for downstream tasks. Clustering S3M representations…