1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.SD2026
A Human-Inspired Decoupled Architecture for Efficient Audio Representation Learning
Harunori Kawano, Takeshi Sasaki
While self-supervised learning (SSL) has revolutionized audio representation, the excessive parameterization and quadratic computational cost of standard Transformers limit their d…
cs.SD2023★ 1 cited
An Effective Transformer-based Contextual Model and Temporal Gate Pooling for Speaker Identification
Harunori Kawano, Sota Shimizu
Wav2vec2 has achieved success in applying Transformer architecture and self-supervised learning to speech recognition. Recently, these have come to be used not only for speech reco…