most citedBalalaika: Data-Centric, Prosody-Aware Annotation Pipeline for Russian Speech

1 citations · 1 across the 4 of their papers we have counts for

collaborators

11 papers

cs.LG2026

Think Shallow, Solve Deep: Controlling Recurrent Dynamics for Reliable Test-Time Depth

Ivan Viakhirev, Kirill Borodin, Amirah Almutairi +3

Recurrent-depth reasoners aim to solve harder problems by iterating their update longer at test time, but additional iterations can improve, preserve, or degrade an answer. We show…

cs.SD2026

Training-Free Model Selection and Domain-Aware Score Calibration for First-Shot Anomalous Sound Detection

Grach Mkrtchian

First-shot anomalous sound detection in DCASE Challenge Task 2 must flag anomalies of unseen machine types with a single threshold, without knowing whether a test clip comes from t…

cs.SD2026

When Spoof Detectors Travel: Evaluation Across 66 Languages in the Low-Resource Language Spoofing Corpus

Kirill Borodin, Vasiliy Kudryavtsev, Maxim Maslov +2

We introduce LRLspoof, a large-scale multilingual synthetic-speech corpus for cross-lingual spoof detection, comprising 2,732 hours of audio generated with 24 open-source TTS syste…

cs.LG2026

From Dispersion to Attraction: Spectral Dynamics of Hallucination Across Whisper Model Scales

Ivan Viakhirev, Kirill Borodin, Grach Mkrtchian

Hallucinations in large ASR models present a critical safety risk. In this work, we propose the \textit{Spectral Sensitivity Theorem}, which predicts a phase transition in deep net…

cs.CL20261 cited

Balalaika: Data-Centric, Prosody-Aware Annotation Pipeline for Russian Speech

Kirill Borodin, Nikita Vasiliev, Vasiliy Kudryavtsev +3

We introduce Balalaika, an open-source, data-centric pipeline for processing audio and producing prosody-aware annotations. It combines semantic VAD for context-preserving segmenta…

cs.SD2026

AASIST3: KAN-Enhanced AASIST Speech Deepfake Detection using SSL Features and Additional Regularization for the ASVspoof 2024 Challenge

Kirill Borodin, Vasiliy Kudryavtsev, Dmitrii Korzh +4

Automatic Speaker Verification (ASV) systems, which identify speakers based on their voice characteristics, have numerous applications, such as user authentication in financial tra…