5 papers
Multi-granularity Interactive Attention Framework for Residual Hierarchical Pronunciation Assessment
Hong Han, Hao-Chen Pei, Zhao-Zheng Nie +2
Automatic pronunciation assessment plays a crucial role in computer-assisted pronunciation training systems. Due to the ability to perform multiple pronunciation tasks simultaneous…
Suppressing Gradient Conflict for Generalizable Deepfake Detection
Ming-Hui Liu, Harry Cheng, Xin Luo +1
Robust deepfake detection models must be capable of generalizing to ever-evolving manipulation techniques beyond training data. A promising strategy is to augment the training data…
Learning Real Facial Concepts for Independent Deepfake Detection
Ming-Hui Liu, Harry Cheng, Tianyi Wang +2
Deepfake detection models often struggle with generalization to unseen datasets, manifesting as misclassifying real instances as fake in target domains. This is primarily due to an…
DATA: Multi-Disentanglement based Contrastive Learning for Open-World Semi-Supervised Deepfake Attribution
Ming-Hui Liu, Xiao-Qian Liu, Xin Luo +1
Deepfake attribution (DFA) aims to perform multiclassification on different facial manipulation techniques, thereby mitigating the detrimental effects of forgery content on the soc…
Enhancing HOI Detection with Contextual Cues from Large Vision-Language Models
Yu-Wei Zhan, Fan Liu, Xin Luo +3
Human-Object Interaction (HOI) detection aims at detecting human-object pairs and predicting their interactions. However, conventional HOI detection methods often struggle to fully…