activity
20212025
most citedMore for Less: Non-Intrusive Speech Quality Assessment with Limited Annotations

14 citations · 15 across the 5 of their papers we have counts for

collaborators
Showing cs.SDShow all

5 papers · 1 filter

cs.SD20251 cited

Binaspect -- A Python Library for Binaural Audio Analysis, Visualization & Feature Generation

Dan Barry, Davoud Shariat Panah, Alessandro Ragano +2

We present Binaspect, an open-source Python library for binaural audio analysis, visualization, and feature generation. Binaspect generates interpretable "azimuth maps" by calculat…

cs.SD2025

Binamix -- A Python Library for Generating Binaural Audio Datasets

Dan Barry, Davoud Shariat Panah, Alessandro Ragano +2

The increasing demand for spatial audio in applications such as virtual reality, immersive media, and spatial audio research necessitates robust solutions to generate binaural audi…

cs.SD2024

SCOREQ: Speech Quality Assessment with Contrastive Regression

Alessandro Ragano, Jan Skoglund, Andrew Hines

In this paper, we present SCOREQ, a novel approach for speech quality prediction. SCOREQ is a triplet loss function for contrastive regression that addresses the domain generalisat…

cs.SD2023

NOMAD: Unsupervised Learning of Perceptual Embeddings for Speech Enhancement and Non-matching Reference Audio Quality Assessment

Alessandro Ragano, Jan Skoglund, Andrew Hines

This paper presents NOMAD (Non-Matching Audio Distance), a differentiable perceptual similarity metric that measures the distance of a degraded signal against non-matching referenc…

cs.SD2022

Using Rater and System Metadata to Explain Variance in the VoiceMOS Challenge 2022 Dataset

Michael Chinen, Jan Skoglund, Chandan K A Reddy +2

Non-reference speech quality models are important for a growing number of applications. The VoiceMOS 2022 challenge provided a dataset of synthetic voice conversion and text-to-spe…