39 citations · 45 across the 8 of their papers we have counts for
17 papers · 1 filter
QuAVF: Quality-aware Audio-Visual Fusion for Ego4D Talking to Me Challenge
Hsi-Che Lin, Chien-Yi Wang, Min-Hung Chen +2
This technical report describes our QuAVF@NTU-NVIDIA submission to the Ego4D Talking to Me (TTM) Challenge 2023. Based on the observation from the TTM task and the provided dataset…
A Closer Look at Geometric Temporal Dynamics for Face Anti-Spoofing
Chih-Jung Chang, Yaw-Chern Lee, Shih-Hsuan Yao +4
Face anti-spoofing (FAS) is indispensable for a face recognition system. Many texture-driven countermeasures were developed against presentation attacks (PAs), but the performance…
Causalainer: Causal Explainer for Automatic Video Summarization
Jia-Hong Huang, Chao-Han Huck Yang, Pin-Yu Chen +2
The goal of video summarization is to automatically shorten videos such that it conveys the overall story without losing relevant information. In many application scenarios, improp…
Interaction-Aware Prompting for Zero-Shot Spatio-Temporal Action Detection
Wei-Jhe Huang, Jheng-Hsien Yeh, Min-Hung Chen +2
The goal of spatial-temporal action detection is to determine the time and place where each person's action occurs in a video and classify the corresponding action category. Most o…
Kinship Representation Learning with Face Componential Relation
Weng-Tai Su, Min-Hung Chen, Chien-Yi Wang +2
Kinship recognition aims to determine whether the subjects in two facial images are kin or non-kin, which is an emerging and challenging problem. However, most previous methods foc…
Holistic Interaction Transformer Network for Action Detection
Gueter Josmy Faure, Min-Hung Chen, Shang-Hong Lai
Actions are about how we interact with the environment, including other people, objects, and ourselves. In this paper, we propose a novel multi-modal Holistic Interaction Transform…