activity
20172023
most citedTS-LSTM and Temporal-Inception: Exploiting Spatiotemporal Dynamics for Activity Recognition

39 citations · 45 across the 8 of their papers we have counts for

collaborators
Showing cs.CVShow all

17 papers · 1 filter

cs.CV2023

QuAVF: Quality-aware Audio-Visual Fusion for Ego4D Talking to Me Challenge

Hsi-Che Lin, Chien-Yi Wang, Min-Hung Chen +2

This technical report describes our QuAVF@NTU-NVIDIA submission to the Ego4D Talking to Me (TTM) Challenge 2023. Based on the observation from the TTM task and the provided dataset…

cs.CV2023

A Closer Look at Geometric Temporal Dynamics for Face Anti-Spoofing

Chih-Jung Chang, Yaw-Chern Lee, Shih-Hsuan Yao +4

Face anti-spoofing (FAS) is indispensable for a face recognition system. Many texture-driven countermeasures were developed against presentation attacks (PAs), but the performance…

cs.CV2023

Causalainer: Causal Explainer for Automatic Video Summarization

Jia-Hong Huang, Chao-Han Huck Yang, Pin-Yu Chen +2

The goal of video summarization is to automatically shorten videos such that it conveys the overall story without losing relevant information. In many application scenarios, improp…

cs.CV2023

Interaction-Aware Prompting for Zero-Shot Spatio-Temporal Action Detection

Wei-Jhe Huang, Jheng-Hsien Yeh, Min-Hung Chen +2

The goal of spatial-temporal action detection is to determine the time and place where each person's action occurs in a video and classify the corresponding action category. Most o…

cs.CV2023

Kinship Representation Learning with Face Componential Relation

Weng-Tai Su, Min-Hung Chen, Chien-Yi Wang +2

Kinship recognition aims to determine whether the subjects in two facial images are kin or non-kin, which is an emerging and challenging problem. However, most previous methods foc…

cs.CV20221 cited

Holistic Interaction Transformer Network for Action Detection

Gueter Josmy Faure, Min-Hung Chen, Shang-Hong Lai

Actions are about how we interact with the environment, including other people, objects, and ourselves. In this paper, we propose a novel multi-modal Holistic Interaction Transform…