7 papers
EgoAgent: A Joint Predictive Agent Model in Egocentric Worlds
Lu Chen, Yizhou Wang, Shixiang Tang +6
Learning an agent model that behaves like humans-capable of jointly perceiving the environment, predicting the future, and taking actions from a first-person perspective-is a funda…
Depth Any Video with Scalable Synthetic Data
Honghui Yang, Di Huang, Wei Yin +6
Video depth estimation has long been hindered by the scarcity of consistent and scalable ground truth data, leading to inconsistent and unreliable results. In this paper, we introd…
NeuRodin: A Two-stage Framework for High-Fidelity Neural Surface Reconstruction
Yifan Wang, Di Huang, Weicai Ye +3
Signed Distance Function (SDF)-based volume rendering has demonstrated significant capabilities in surface reconstruction. Although promising, SDF-based methods often fail to captu…
TASeg: Temporal Aggregation Network for LiDAR Semantic Segmentation
Xiaopei Wu, Yuenan Hou, Xiaoshui Huang +8
Training deep models for LiDAR semantic segmentation is challenging due to the inherent sparsity of point clouds. Utilizing temporal data is a natural remedy against the sparsity p…
PredBench: Benchmarking Spatio-Temporal Prediction across Diverse Disciplines
ZiDong Wang, Zeyu Lu, Di Huang +4
In this paper, we introduce PredBench, a benchmark tailored for the holistic evaluation of spatio-temporal prediction networks. Despite significant progress in this field, there re…
Holistic-Motion2D: Scalable Whole-body Human Motion Generation in 2D Space
Yuan Wang, Zhao Wang, Junhao Gong +8
In this paper, we introduce a novel path to human motion generation by focusing on 2D space. Traditional methods have primarily generated human motions in 3D, wh…