4 citations · 6 across the 3 of their papers we have counts for
5 papers · 1 filter
YouTube-Occ: Learning Indoor 3D Semantic Occupancy Prediction from YouTube Videos
Haoming Chen, Lichen Yuan, TianFang Sun +5
3D semantic occupancy prediction is crucial for fine-grained scene understanding, yet its advancement in privacy-sensitive indoor environments is fundamentally hindered by the scar…
Learning semantical dynamics and spatiotemporal collaboration for human pose estimation in video
Runyang Feng, Haoming Chen
Temporal modeling and spatio-temporal collaboration are pivotal techniques for video-based human pose estimation. Most state-of-the-art methods adopt optical flow or temporal diffe…
Building a Strong Pre-Training Baseline for Universal 3D Large-Scale Perception
Haoming Chen, Zhizhong Zhang, Yanyun Qu +3
An effective pre-training framework with universal 3D representations is extremely desired in perceiving large-scale dynamic scenes. However, establishing such an ideal framework t…
2D Human Pose Estimation: A Survey
Haoming Chen, Runyang Feng, Sifan Wu +3
Human pose estimation aims at localizing human anatomical keypoints or body parts in the input data (e.g., images, videos, or signals). It forms a crucial component in enabling mac…
Temporal Feature Alignment and Mutual Information Maximization for Video-Based Human Pose Estimation
Zhenguang Liu, Runyang Feng, Haoming Chen +4
Multi-frame human pose estimation has long been a compelling and fundamental problem in computer vision. This task is challenging due to fast motion and pose occlusion that frequen…