4 papers · 1 filter
Affinity Contrastive Learning for Skeleton-based Human Activity Understanding
Hongda Liu, Yunfan Liu, Min Ren +3
In skeleton-based human activity understanding, existing methods often adopt the contrastive learning paradigm to construct a discriminative feature space. However, many of these a…
StableMotion: One-Step Motion Estimation with Diffusion Prior
Ziyi Wang, Haipeng Li, Lin Sui +5
We present StableMotion, a novel framework that leverages geometric and content priors from pretrained large-scale image diffusion models for motion estimation in single-image rect…
TimeLoc: A Unified End-to-End Framework for Precise Timestamp Localization in Long Videos
Chen-Lin Zhang, Lin Sui, Shuming Liu +3
Temporal localization in untrimmed videos, which aims to identify specific timestamps, is crucial for video understanding but remains challenging. This task encompasses several sub…
Harnessing Temporal Causality for Advanced Temporal Action Detection
Shuming Liu, Lin Sui, Chen-Lin Zhang +3
As a fundamental task in long-form video understanding, temporal action detection (TAD) aims to capture inherent temporal relations in untrimmed videos and identify candidate actio…