1 paper
Soonshin Seo, Ji-Hwan Kim
In general, a self-attention mechanism has been applied for speaker embedding encoding. Previous studies focused on training the self-attention in a high-level layer, such as the l…