activity
20242026
collaborators

5 papers

cs.CL2026

Multi-granularity Interactive Attention Framework for Residual Hierarchical Pronunciation Assessment

Hong Han, Hao-Chen Pei, Zhao-Zheng Nie +2

Automatic pronunciation assessment plays a crucial role in computer-assisted pronunciation training systems. Due to the ability to perform multiple pronunciation tasks simultaneous…

cs.CV2025

Suppressing Gradient Conflict for Generalizable Deepfake Detection

Ming-Hui Liu, Harry Cheng, Xin Luo +1

Robust deepfake detection models must be capable of generalizing to ever-evolving manipulation techniques beyond training data. A promising strategy is to augment the training data…

cs.CV2025

Learning Real Facial Concepts for Independent Deepfake Detection

Ming-Hui Liu, Harry Cheng, Tianyi Wang +2

Deepfake detection models often struggle with generalization to unseen datasets, manifesting as misclassifying real instances as fake in target domains. This is primarily due to an…

cs.CV2025

DATA: Multi-Disentanglement based Contrastive Learning for Open-World Semi-Supervised Deepfake Attribution

Ming-Hui Liu, Xiao-Qian Liu, Xin Luo +1

Deepfake attribution (DFA) aims to perform multiclassification on different facial manipulation techniques, thereby mitigating the detrimental effects of forgery content on the soc…

cs.CV2024

Enhancing HOI Detection with Contextual Cues from Large Vision-Language Models

Yu-Wei Zhan, Fan Liu, Xin Luo +3

Human-Object Interaction (HOI) detection aims at detecting human-object pairs and predicting their interactions. However, conventional HOI detection methods often struggle to fully…