collaborators

7 papers

cs.CL2026

Emotion Across Speech and Faces: Shared Affective Mechanisms in Multimodal Foundation Models

Xiutian Zhao, Luqi Sun, Björn Schuller +1

Modern multimodal foundation models (MFMs) have made rapid progress on tasks requiring integrated perception across speech, vision, and language, including emotion recognition. How…

cs.CL2026

Multilingual Emotion Neurons in Large Audio-Language Models

Xiutian Zhao, Philipp Koehn, Björn Schuller +1

Emotion is central to human communication, and its expression varies across languages. Large audio-language models (LALMs) achieve strong performance on multilingual speech tasks,…

cs.SD2026

The Affective Bridge: Preserving Speech Representations while Enhancing Deepfake Detection vian emotional Constraints

Yupei Li, Chenyang Lyu, Longyue Wang +4

Speech deepfake detection (DFD) has benefited from diverse acoustic and semantic speech representations, many of which encode valuable speech information and are costly to train. P…

cs.CL2026

Feature-Augmented Transformers for Robust AI-Text Detection Across Domains and Generators

Mohamed Mady, Johannes Reschke, Björn Schuller

AI-generated text is nowadays produced at scale across domains and heterogeneous generation pipelines, making robustness to distribution shift a central requirement for supervised…

eess.AS2026

Explainable Speech Emotion Recognition: Weighted Attribute Fairness to Model Demographic Contributions to Social Bias

Tomisin Ogunnubi, Yupei Li, Björn Schuller

Speech Emotion Recognition (SER) systems have growing applications in sensitive domains such as mental health and education, where biased predictions can cause harm. Traditional fa…

cs.CL2026

Neuron-Level Emotion Control in Speech-Generative Large Audio-Language Models

Xiutian Zhao, Ismail Rasim Ulgen, Philipp Koehn +2

Large audio-language models (LALMs) can produce expressive speech, yet reliable emotion control remains elusive: conversions often miss the target affect and may degrade linguistic…