From the 1 of 6 linked papers with an AI index.
6 papers
CF-Net: Conflict Fusion with Speaker Normalisation and Certainty Weighting for Ambivalence/Hesitancy Recognition
Tung Hung Bui, Hong Hai Nguyen, Van Thong Huynh
The paper introduces CF-Net, a multimodal deep network that detects ambivalence and hesitancy in videos by fusing visual, audio, and transcript features while normalizing per speak…
Strength-Parity Ensembling with Parameter-Isolated Experts for Multi-Task Affect Recognition
Hong Hai Nguyen, Tung Hung Bui, Van Thong Huynh
Leading entries on the multi-task track of the 11th ABAW challenge rely on heavy ensembling, yet which member is worth adding to an already strong ensemble is rarely made explicit.…
A Shared Latent for Partially-Labeled Multi-Task Facial Affect Recognition
Hong Hai Nguyen, Sy Phan Van, Soo-Hyung Kim +1
Facial affect in the wild is naturally multi-task: valence-arousal, discrete expressions, and facial action units describe the same face. Yet real corpora annotate these tasks only…
Faithful Action-unit Causal Reasoning for Counterfactually Faithful Emotion Explanations
Van Thong Huynh, Hong Hai Nguyen, Thuy Pham +2
Multimodal models can name the action units (AUs) behind a facial emotion, but their AU->emotion rationales are typically plausible rather than faithful: nothing forces the AUs a m…
The Circumplex Degeneracy Behind the Rare-Class Limit in Affect Recognition
Van Thong Huynh, Hong Hai Nguyen, Soo-Hyung Kim
In-the-wild expression recognition persistently fails on a few rare emotions, and the standard explanation is class imbalance. Through a controlled multi-task study on two benchmar…
Lightweight Models for Emotional Analysis in Video
Quoc-Tien Nguyen, Hong-Hai Nguyen, Van-Thong Huynh
In this study, we present an approach for efficient spatiotemporal feature extraction using MobileNetV4 and a multi-scale 3D MLP-Mixer-based temporal aggregation module. MobileNetV…