activity
20222024
most citedRedundancy Reduction Twins Network: A Training framework for Multi-output Emotion Regression

4 citations · 10 across the 5 of their papers we have counts for

collaborators

5 papers

cs.SD20241 cited

STAA-Net: A Sparse and Transferable Adversarial Attack for Speech Emotion Recognition

Yi Chang, Zhao Ren, Zixing Zhang +6

Speech contains rich information on the emotions of humans, and Speech Emotion Recognition (SER) has been an important topic in the area of human-computer interaction. The robustne…

cs.SD20231 cited

U-DiT TTS: U-Diffusion Vision Transformer for Text-to-Speech

Xin Jing, Yi Chang, Zijiang Yang +3

Deep learning has led to considerable advances in text-to-speech synthesis. Most recently, the adoption of Score-based Generative Models (SGMs), also known as Diffusion Probabilist…

cs.SD2023

HEAR4Health: A blueprint for making computer audition a staple of modern healthcare

Andreas Triantafyllopoulos, Alexander Kathan, Alice Baird +20

Recent years have seen a rapid increase in digital medicine research in an attempt to transform traditional healthcare systems to their modern, intelligent, and versatile equivalen…

cs.SD20224 cited

Redundancy Reduction Twins Network: A Training framework for Multi-output Emotion Regression

Xin Jing, Meishu Song, Andreas Triantafyllopoulos +2

In this paper, we propose the Redundancy Reduction Twins Network (RRTN), a redundancy reduction training framework that minimizes redundancy by measuring the cross-correlation matr…

cs.SD20224 cited

Dynamic Restrained Uncertainty Weighting Loss for Multitask Learning of Vocal Expression

Meishu Song, Zijiang Yang, Andreas Triantafyllopoulos +6

We propose a novel Dynamic Restrained Uncertainty Weighting Loss to experimentally handle the problem of balancing the contributions of multiple tasks on the ICML ExVo 2022 Challen…