works on

From the 1 of 6 linked papers with an AI index.

collaborators

6 papers

cs.CV2026

CF-Net: Conflict Fusion with Speaker Normalisation and Certainty Weighting for Ambivalence/Hesitancy Recognition

Tung Hung Bui, Hong Hai Nguyen, Van Thong Huynh

The paper introduces CF-Net, a multimodal deep network that detects ambivalence and hesitancy in videos by fusing visual, audio, and transcript features while normalizing per speak…

cs.CV2026

Strength-Parity Ensembling with Parameter-Isolated Experts for Multi-Task Affect Recognition

Hong Hai Nguyen, Tung Hung Bui, Van Thong Huynh

Leading entries on the multi-task track of the 11th ABAW challenge rely on heavy ensembling, yet which member is worth adding to an already strong ensemble is rarely made explicit.…

cs.CV2026

A Shared Latent for Partially-Labeled Multi-Task Facial Affect Recognition

Hong Hai Nguyen, Sy Phan Van, Soo-Hyung Kim +1

Facial affect in the wild is naturally multi-task: valence-arousal, discrete expressions, and facial action units describe the same face. Yet real corpora annotate these tasks only…

cs.CV2026

Faithful Action-unit Causal Reasoning for Counterfactually Faithful Emotion Explanations

Van Thong Huynh, Hong Hai Nguyen, Thuy Pham +2

Multimodal models can name the action units (AUs) behind a facial emotion, but their AU->emotion rationales are typically plausible rather than faithful: nothing forces the AUs a m…

cs.CV2026

The Circumplex Degeneracy Behind the Rare-Class Limit in Affect Recognition

Van Thong Huynh, Hong Hai Nguyen, Soo-Hyung Kim

In-the-wild expression recognition persistently fails on a few rare emotions, and the standard explanation is class imbalance. Through a controlled multi-task study on two benchmar…

cs.CV2025

Lightweight Models for Emotional Analysis in Video

Quoc-Tien Nguyen, Hong-Hai Nguyen, Van-Thong Huynh

In this study, we present an approach for efficient spatiotemporal feature extraction using MobileNetV4 and a multi-scale 3D MLP-Mixer-based temporal aggregation module. MobileNetV…