papers

Publications (22)

cs.CV2020

Hybrid Dynamic-static Context-aware Attention Network for Action Assessment in Long Videos

Ling-An Zeng, Fa-Ting Hong, Wei-Shi Zheng +4

The objective of action quality assessment is to score sports videos. However, most existing works focus only on video dynamic information (i.e., motion information) but ignore the…

cs.CV2020

MINI-Net: Multiple Instance Ranking Network for Video Highlight Detection

Fa-Ting Hong, Xuanteng Huang, Wei-Hong Li +1

We address the weakly supervised video highlight detection problem for learning to detect segments that are more attractive in training videos given their video event label but wit…

cs.CV2025

Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation

Shuling Zhao, Fa-Ting Hong, Xiaoshui Huang +1

Talking head video generation aims to generate a realistic talking head video that preserves the person's identity from a source image and the motion from a driving video. Despite…

cs.CV2024

Learning Online Scale Transformation for Talking Head Video Generation

Fa-Ting Hong, Dan Xu

One-shot talking head video generation uses a source image and driving video to create a synthetic video where the source person's facial movements imitate those of the driving vid…

cs.CV2023

Implicit Identity Representation Conditioned Memory Compensation Network for Talking Head video Generation

Fa-Ting Hong, Dan Xu

Talking head video generation aims to animate a human face in a still image with dynamic poses and expressions using motion information derived from a target-driving video, while m…

cs.CV2026

PhysiGen: Integrating Collision-Aware Physical Constraints for High-Fidelity Human-Human Interaction Generation

Nan Lei, Yuan-Ming Li, Ling-An Zeng +5

Despite substantial progress in text-driven 3D human motion synthesis, generating realistic multi-person interaction sequences remains challenging. Notably, body inter-penetration…