From the 2 of 11 linked papers with an AI index.
11 papers
Large Audio Language Models for Spoofing-Aware Speaker Verification
Sofya Savelyeva, Mariia Perunova, Evgeny Kushnir +3
The paper investigates the use of large audio language models for spoofing‑aware speaker verification, evaluating zero‑shot, supervised, reasoning‑oriented, and reinforcement‑learn…
Towards Robust Speech Deepfake Detection via Human-Inspired Reasoning
Artem Dvirniak, Evgeny Kushnir, Dmitrii Tarasov +5
The paper introduces HIR‑SDD, a speech deepfake detection framework that leverages large audio language models and chain‑of‑thought reasoning from a human‑annotated dataset to impr…
AASIST3: KAN-Enhanced AASIST Speech Deepfake Detection using SSL Features and Additional Regularization for the ASVspoof 2024 Challenge
Kirill Borodin, Vasiliy Kudryavtsev, Dmitrii Korzh +4
Automatic Speaker Verification (ASV) systems, which identify speakers based on their voice characteristics, have numerous applications, such as user authentication in financial tra…
LLM-Guided Prompt Evolution for Password Guessing
Vladimir A. Mazin, Mikhail A. Zorin, Dmitrii S. Korzh +3
Passwords still remain a dominant authentication method, yet their security is routinely subverted by predictable user choices and large-scale credential leaks. Automated password…
Probabilistic Verification of Voice Anti-Spoofing Models
Evgeny Kushnir, Alexandr Kozodaev, Dmitrii Korzh +3
Recent advances in generative models have amplified the risk of malicious misuse of speech synthesis technologies, enabling adversaries to impersonate target speakers and access se…
Speech-to-LaTeX: New Models and Datasets for Converting Spoken Equations and Sentences
Dmitrii Korzh, Dmitrii Tarasov, Artyom Iudin +6
Conversion of spoken mathematical expressions is a challenging task that involves transcribing speech into a strictly structured symbolic representation while addressing the ambigu…