3 papers
eess.AS2025
On-device Streaming Discrete Speech Units
Kwanghee Choi, Masao Someki, Emma Strubell +1
Discrete speech units (DSUs) are derived from clustering the features of self-supervised speech models (S3Ms). DSUs offer significant advantages for on-device streaming speech appl…
eess.AS2025
Context-Driven Dynamic Pruning for Large Speech Foundation Models
Masao Someki, Shikhar Bharadwaj, Atharva Anand Joshi +7
Speech foundation models achieve strong generalization across languages and acoustic conditions, but require significant computational resources for inference. In the context of sp…
eess.AS2022
ESPnet-ONNX: Bridging a Gap Between Research and Production
Masao Someki, Yosuke Higuchi, Tomoki Hayashi +1
In the field of deep learning, researchers often focus on inventing novel neural network models and improving benchmarks. In contrast, application developers are interested in maki…