6 papers · 1 filter
Investigating the impact of stereo processing -- a study for extending the Open Dataset of Audio Quality (ODAQ)
Sascha Dick, Christoph Thompson, Chih-Wei Wu +3
In this paper, we present an initial study for extending Open Dataset of Audio Quality (ODAQ) towards the impact of stereo processing. Monaural artifacts from ODAQ were adapted in…
Navigating PESQ: Up-to-Date Versions and Open Implementations
Matteo Torcoli, Mhd Modar Halimeh, Emanuël A. P. Habets
Perceptual Evaluation of Speech Quality (PESQ) is an objective quality measure that remains widely used despite its withdrawal by the International Telecommunication Union (ITU). P…
Expanding and Analyzing ODAQ -- the Open Dataset of Audio Quality
Sascha Dick, Christoph Thompson, Chih-Wei Wu +4
The Open Dataset of Audio Quality (ODAQ) was recently introduced to address the scarcity of openly available audio datasets with corresponding subjective quality scores. The datase…
On the Relation Between Speech Quality and Quantized Latent Representations of Neural Codecs
Mhd Modar Halimeh, Matteo Torcoli, Philipp Grundhuber +1
Neural audio signal codecs have attracted significant attention in recent years. In essence, the impressive low bitrate achieved by such encoders is enabled by learning an abstract…
ConcateNet: Dialogue Separation Using Local And Global Feature Concatenation
Mhd Modar Halimeh, Matteo Torcoli, Emanuël Habets
Dialogue separation involves isolating a dialogue signal from a mixture, such as a movie or a TV program. This can be a necessary step to enable dialogue enhancement for broadcast-…
Speech Loudness in Broadcasting and Streaming
Matteo Torcoli, Mhd Modar Halimeh, Thomas Leitz +6
The introduction and regulation of loudness in broadcasting and streaming brought clear benefits to the audience, e.g., a level of uniformity across programs and channels. Yet, spe…