2 papers
eess.AS2025
Discrete Speech Unit Extraction via Independent Component Analysis
Tomohiko Nakamura, Kwanghee Choi, Keigo Hojo +3
Self-supervised speech models (S3Ms) have become a common tool for the speech processing community, leveraging representations for downstream tasks. Clustering S3M representations…
eess.AS2024
Neural Blind Source Separation and Diarization for Distant Speech Recognition
Yoshiaki Bando, Tomohiko Nakamura, Shinji Watanabe
This paper presents a neural method for distant speech recognition (DSR) that jointly separates and diarizes speech mixtures without supervision by isolated signals. A standard sep…