activity
20242026
collaborators
Showing cs.CVShow all

15 papers · 1 filter

cs.CV2026

MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding

Yuan Wang, Di Huang, Yaqi Zhang +5

Generating lifelike human motions from descriptive texts has experienced remarkable research focus in the recent years, propelled by the emerging requirements of digital humans.Des…

cs.CV2025

Depth Any Video with Scalable Synthetic Data

Honghui Yang, Di Huang, Wei Yin +6

Video depth estimation has long been hindered by the scarcity of consistent and scalable ground truth data, leading to inconsistent and unreliable results. In this paper, we introd…

cs.CV2024

NeuRodin: A Two-stage Framework for High-Fidelity Neural Surface Reconstruction

Yifan Wang, Di Huang, Weicai Ye +3

Signed Distance Function (SDF)-based volume rendering has demonstrated significant capabilities in surface reconstruction. Although promising, SDF-based methods often fail to captu…

cs.CV2024

DATAP-SfM: Dynamic-Aware Tracking Any Point for Robust Structure from Motion in the Wild

Weicai Ye, Xinyu Chen, Ruohao Zhan +7

This paper proposes a concise, elegant, and robust pipeline to estimate smooth camera trajectories and obtain dense point clouds for casual videos in the wild. Traditional framewor…

cs.CV2024

DiffPano: Scalable and Consistent Text to Panorama Generation with Spherical Epipolar-Aware Diffusion

Weicai Ye, Chenhao Ji, Zheng Chen +7

Diffusion-based methods have achieved remarkable achievements in 2D image or 3D object generation, however, the generation of 3D scenes and even images remains constr…

cs.CV2024

Where Am I and What Will I See: An Auto-Regressive Model for Spatial Localization and View Prediction

Junyi Chen, Di Huang, Weicai Ye +2

Spatial intelligence is the ability of a machine to perceive, reason, and act in three dimensions within space and time. Recent advancements in large-scale auto-regressive models h…