25 citations · 61 across the 14 of their papers we have counts for
13 papers · 1 filter
Social-MAE: Social Masked Autoencoder for Multi-person Motion Representation Learning
Mahsa Ehsanpour, Ian Reid, Hamid Rezatofighi
For a complete comprehension of multi-person scenes, it is essential to go beyond basic tasks like detection and tracking. Higher-level tasks, such as understanding the interaction…
Physically Plausible 3D Human-Scene Reconstruction from Monocular RGB Image using an Adversarial Learning Approach
Sandika Biswas, Kejie Li, Biplab Banerjee +2
Holistic 3D human-scene reconstruction is a crucial and emerging research area in robot perception. A key challenge in holistic 3D human-scene reconstruction is to generate a physi…
ActiveRMAP: Radiance Field for Active Mapping And Planning
Huangying Zhan, Jiyang Zheng, Yi Xu +2
A high-quality 3D reconstruction of a scene from a collection of 2D images can be achieved through offline/online mapping methods. In this paper, we explore active mapping from the…
SoMoFormer: Multi-Person Pose Forecasting with Transformers
Edward Vendrow, Satyajit Kumar, Ehsan Adeli +1
Human pose forecasting is a challenging problem involving complex human body motion and posture dynamics. In cases that there are multiple people in the environment, one's motion m…
ODAM: Object Detection, Association, and Mapping using Posed RGB Video
Kejie Li, Daniel DeTone, Steven Chen +6
Localizing objects and estimating their extent in 3D is an important step towards high-level 3D scene understanding, which has many applications in Augmented Reality and Robotics.…
JRDB-Act: A Large-scale Dataset for Spatio-temporal Action, Social Group and Activity Detection
Mahsa Ehsanpour, Fatemeh Saleh, Silvio Savarese +2
The availability of large-scale video action understanding datasets has facilitated advances in the interpretation of visual scenes containing people. However, learning to recognis…