Showing eess.ASShow all
3 papers · 1 filter
eess.AS2026
ReDimNet2: Scaling Speaker Verification via Time-Pooled Dimension Reshaping
Ivan Yakovlev, Anton Okhotnikov
We present ReDimNet2, an improved neural network architecture for extracting utterance-level speaker representations that builds upon the ReDimNet dimension-reshaping framework. Th…
eess.AS2024
Study on Inter and Intra Speaker Variability in Speaker Recognition
Anton Okhotnikov, Nikita Torgashov, Ivan Yakovlev +2
Optimization of a trade-off between the number of speakers and their temporal variability (or session diversity) is crucial for the development of a speaker recognition system toge…
eess.AS2024
Reshape Dimensions Network for Speaker Recognition
Ivan Yakovlev, Rostislav Makarov, Andrei Balykin +3
In this paper, we present Reshape Dimensions Network (ReDimNet), a novel neural network architecture for extracting utterance-level speaker representations. Our approach leverages…