15 citations · 16 across the 2 of their papers we have counts for
2 papers
cs.SD2021★ 15 cited
DECAR: Deep Clustering for learning general-purpose Audio Representations
Sreyan Ghosh, Sandesh V Katta, Ashish Seth +1
We introduce DECAR, a self-supervised pre-training approach for learning general-purpose audio representations. Our system is based on clustering: it utilizes an offline clustering…
eess.AS2020★ 1 cited
S-vectors and TESA: Speaker Embeddings and a Speaker Authenticator Based on Transformer Encoder
N J Metilda Sagaya Mary, S Umesh, Sandesh V Katta
One of the most popular speaker embeddings is x-vectors, which are obtained from an architecture that gradually builds a larger temporal context with layers. In this paper, we prop…