Publications (11)
Adversarial Cross-Domain Action Recognition with Co-Attention
Boxiao Pan, Zhangjie Cao, Ehsan Adeli +1
Action recognition has been a widely studied topic with a heavy focus on supervised learning involving sufficient labeled videos. However, the problem of cross-domain action recogn…
Spatio-Temporal Graph for Video Captioning with Knowledge Distillation
Boxiao Pan, Haoye Cai, De-An Huang +4
Video captioning is a challenging task that requires a deep understanding of visual scenes. State-of-the-art methods generate captions using either scene-level or object-level info…
MultiPhys: Multi-Person Physics-aware 3D Motion Estimation
Nicolas Ugrinovic, Boxiao Pan, Georgios Pavlakos +5
We introduce MultiPhys, a method designed for recovering multi-person motion from monocular videos. Our focus lies in capturing coherent spatial placement between pairs of individu…
COPILOT: Human-Environment Collision Prediction and Localization from Egocentric Videos
Boxiao Pan, Bokui Shen, Davis Rempe +4
The ability to forecast human-environment collisions from egocentric observations is vital to enable collision avoidance in applications such as VR, AR, and wearable assistive robo…
PartNeRF: Generating Part-Aware Editable 3D Shapes without 3D Supervision
Konstantinos Tertikas, Despoina Paschalidou, Boxiao Pan +5
Impressive progress in generative models and implicit representations gave rise to methods that can generate 3D shapes of high quality. However, being able to locally control and e…
LookOut: Real-World Humanoid Egocentric Navigation
Boxiao Pan, Adam W. Harley, C. Karen Liu +1
The ability to predict collision-free future trajectories from egocentric observations is crucial in applications such as humanoid robotics, VR / AR, and assistive navigation. In t…