activity
20242026
collaborators

43 papers

cs.CV2026

SM4RT: Learning Structured Motion Geometry for 4D Reconstruction

Shing Ho J. Lin, Wenzhao Zheng, Dong Zhuo +3

Geometry Foundation Models (GFMs) have substantially advanced monocular 3D reconstruction, yet extending this capability to 4D dynamic understanding remains a fundamental challenge…

cs.CV2026

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation

Yanran Zhang, Ziyi Wang, Wenzhao Zheng +3

Generating interactive and dynamic 4D scenes from a single static image remains a core challenge. Most existing generate-then-reconstruct and reconstruct-then-generate methods deco…

cs.CV2026

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation

Weijia Dou, Wenzhao Zheng, Weiliang Chen +3

Recent generative models can produce high-fidelity videos, yet they often exhibit 3D spatial geometric inconsistencies. Existing evaluation methods fail to accurately characterize…

cs.CV2026

SceneCompleter: Dense 3D Scene Completion for Generative Novel View Synthesis

Weiliang Chen, Jiayi Bi, Yuanhui Huang +2

Generative models have shown great promise for novel view synthesis (NVS) by leveraging strong image generation priors. However, existing approaches typically follow a 2D inpaintin…

cs.CV2026

DVGT: Driving Visual Geometry Transformer

Sicheng Zuo, Zixun Xie, Wenzhao Zheng +6

Perceiving and reconstructing 3D scene geometry from visual inputs is crucial for autonomous driving. However, there still lacks a driving-targeted dense geometry perception model…

cs.RO2026

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation

Wenxuan Guo, Xiuwei Xu, Yichen Liu +7

Vision-and-Language Navigation (VLN) requires an agent to ground language instructions to its own movement within a visual environment. While state-of-the-art methods leverage the…