4 citations · 10 across the 5 of their papers we have counts for
5 papers
STAA-Net: A Sparse and Transferable Adversarial Attack for Speech Emotion Recognition
Yi Chang, Zhao Ren, Zixing Zhang +6
Speech contains rich information on the emotions of humans, and Speech Emotion Recognition (SER) has been an important topic in the area of human-computer interaction. The robustne…
U-DiT TTS: U-Diffusion Vision Transformer for Text-to-Speech
Xin Jing, Yi Chang, Zijiang Yang +3
Deep learning has led to considerable advances in text-to-speech synthesis. Most recently, the adoption of Score-based Generative Models (SGMs), also known as Diffusion Probabilist…
HEAR4Health: A blueprint for making computer audition a staple of modern healthcare
Andreas Triantafyllopoulos, Alexander Kathan, Alice Baird +20
Recent years have seen a rapid increase in digital medicine research in an attempt to transform traditional healthcare systems to their modern, intelligent, and versatile equivalen…
Redundancy Reduction Twins Network: A Training framework for Multi-output Emotion Regression
Xin Jing, Meishu Song, Andreas Triantafyllopoulos +2
In this paper, we propose the Redundancy Reduction Twins Network (RRTN), a redundancy reduction training framework that minimizes redundancy by measuring the cross-correlation matr…
Dynamic Restrained Uncertainty Weighting Loss for Multitask Learning of Vocal Expression
Meishu Song, Zijiang Yang, Andreas Triantafyllopoulos +6
We propose a novel Dynamic Restrained Uncertainty Weighting Loss to experimentally handle the problem of balancing the contributions of multiple tasks on the ICML ExVo 2022 Challen…