3 papers
cs.CV2025
Cameras as Relative Positional Encoding
Ruilong Li, Brent Yi, Junchen Liu +3
Transformers are increasingly prevalent for multi-view computer vision tasks, where geometric relationships between viewpoints are critical for 3D perception. To leverage these rel…
cs.CV2025
Shape of Motion: 4D Reconstruction from a Single Video
Qianqian Wang, Vickie Ye, Hang Gao +4
Monocular dynamic reconstruction is a challenging and long-standing vision problem due to the highly ill-posed nature of the task. Existing approaches depend on templates, are effe…
cs.CV2024
SOAR: Self-Occluded Avatar Recovery from a Single Video In the Wild
Zhuoyang Pan, Angjoo Kanazawa, Hang Gao
Self-occlusion is common when capturing people in the wild, where the performer do not follow predefined motion scripts. This challenges existing monocular human reconstruction sys…