2 citations · 3 across the 5 of their papers we have counts for
5 papers
CL-UZH submission to the NIST SRE 2024 Speaker Recognition Evaluation
Aref Farhadipour, Shiran Liu, Masoumeh Chapariniya +4
The CL-UZH team submitted one system each for the fixed and open conditions of the NIST SRE 2024 challenge. For the closed-set condition, results for the audio-only trials were ach…
Beyond Appearance: Transformer-based Person Identification from Conversational Dynamics
Masoumeh Chapariniya, Teodora Vukovic, Sarah Ebling +1
This paper investigates the performance of transformer-based architectures for person identification in natural, face-to-face conversation scenario. We implement and evaluate a two…
State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition
Aref Farhadipour, Homayoon Beigi, Volker Dellwo +1
Whispered speech recognition presents significant challenges for conventional automatic speech recognition systems, particularly when combined with dialect variation. However, util…
Comparative Analysis of Modality Fusion Approaches for Audio-Visual Person Identification and Verification
Aref Farhadipour, Masoumeh Chapariniya, Teodora Vukovic +1
Multimodal learning involves integrating information from various modalities to enhance learning and comprehension. We compare three modality fusion strategies in person identifica…
Reorganization of the auditory-perceptual space across the human vocal range
Daniel Friedrichs, Volker Dellwo
We analyzed the auditory-perceptual space across a substantial portion of the human vocal range (220-1046 Hz) using multidimensional scaling analysis of cochlea-scaled spectra from…