2 papers
cs.CV2026
ReBaR: Reference-Based Reasoning for Robust Pose Estimation from Monocular Images
Yongkang Cheng, Mingjiang Liang, Jifeng Ning +3
R}easoning for Robust Human Pose and Shape Estimation), designed to estimate human body shape and pose from single-view images. ReBaR effectively addresses the challenges of occlus…
cs.CV2025
DropMAE: Learning Representations via Masked Autoencoders with Spatial-Attention Dropout for Temporal Matching Tasks
Qiangqiang Wu, Tianyu Yang, Ziquan Liu +3
This paper studies masked autoencoder (MAE) video pre-training for various temporal matching-based downstream tasks, i.e., object-level tracking tasks including video object tracki…