activity
20222025
most citedAdvancing Audio Emotion and Intent Recognition with Large Pre-Trained Models and Bayesian Inference

5 citations · 6 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CL2025

Multi-Teacher Language-Aware Knowledge Distillation for Multilingual Speech Emotion Recognition

Mehedi Hasan Bijoy, Dejan Porjazovski, Tamás Grósz +1

Speech Emotion Recognition (SER) is crucial for improving human-computer interaction. Despite strides in monolingual SER, extending them to build a multilingual system remains chal…

cs.CL2024★ 1 cited

Out-of-distribution generalisation in spoken language understanding

Dejan Porjazovski, Anssi Moisio, Mikko Kurimo

Test data is said to be out-of-distribution (OOD) when it unexpectedly differs from the training data, a common challenge in real-world use cases of machine learning. Although OOD…

eess.AS2023★ 5 cited

Advancing Audio Emotion and Intent Recognition with Large Pre-Trained Models and Bayesian Inference

Dejan Porjazovski, Yaroslav Getman, Tamás Grósz +1

Large pre-trained models are essential in paralinguistic systems, demonstrating effectiveness in tasks like emotion recognition and stuttering detection. In this paper, we employ l…

eess.AS2023

Topic Identification For Spontaneous Speech: Enriching Audio Features With Embedded Linguistic Information

Dejan Porjazovski, Tamás Grósz, Mikko Kurimo

Traditional topic identification solutions from audio rely on an automatic speech recognition system (ASR) to produce transcripts used as input to a text-based model. These approac…

cs.CL2022

Lahjoita puhetta -- a large-scale corpus of spoken Finnish with some benchmarks

Anssi Moisio, Dejan Porjazovski, Aku Rouhe +5

The Donate Speech campaign has so far succeeded in gathering approximately 3600 hours of ordinary, colloquial Finnish speech into the Lahjoita puhetta (Donate Speech) corpus. The c…