5 papers
LEGO: Leveled Language Gaussian Splatting
Yuning Peng, Haiping Wang, Yuan Liu +3
We introduce LEGO for advanced open-vocabulary scene understanding. Beyond basic concept recognition, its core innovation lies in capturing the intrinsic semantic hierarchies withi…
SpatialLLM: From Multi-modality Data to Urban Spatial Intelligence
Jiabin Chen, Haiping Wang, Jinpeng Li +3
We propose SpatialLLM, a novel approach advancing spatial intelligence tasks in complex urban scenes. Unlike previous methods requiring geographic analysis tools or domain expertis…
GAGS: Granularity-Aware Feature Distillation for Language Gaussian Splatting
Yuning Peng, Haiping Wang, Yuan Liu +3
3D open-vocabulary scene understanding, which accurately perceives complex semantic properties of objects in space, has gained significant attention in recent years. In this paper,…
Align3R: Aligned Monocular Depth Estimation for Dynamic Videos
Jiahao Lu, Tianyu Huang, Peng Li +7
Recent developments in monocular depth estimation methods enable high-quality depth estimation of single-view images but fail to estimate consistent video depth across different fr…
VistaDream: Sampling multiview consistent images for single-view scene reconstruction
Haiping Wang, Yuan Liu, Ziwei Liu +3
In this paper, we propose VistaDream a novel framework to reconstruct a 3D scene from a single-view image. Recent diffusion models enable generating high-quality novel-view images…