works on

From the 1 of 10 linked papers with an AI index.

collaborators

10 papers

cs.CV2026

ArtiMo: Agent-Driven Articulated Mesh Animation

Chunyu Zou, Peng Dai, Yi-Hua Huang +4

Animating articulated 3D meshes via text requires satisfying strict kinematic constraints, modeling causal interactions between parts, and achieving instruction fidelity. Due to th…

cs.CV2026

FoundationGeo: Learning Spatial Pixel-Wise Fields for Monocular Metric Geometry

Muxin Liu, Xiaoyang Lyu, Tianhe Ren +7

FoundationGeo is a two‑stage framework that first learns an affine‑invariant geometry model from a large multi‑domain dataset, then refines metric depth using lightweight pixel‑wis…

cs.CV2026

S^2VG: 3D Stereoscopic and Spatial Video Generation via Denoising Frame Matrix

Peng Dai, Feitong Tan, Qiangeng Xu +6

While video generation models excel at producing high-quality monocular videos, generating 3D stereoscopic and spatial videos for immersive applications remains an underexplored ch…

cs.CV2026

Stabilizing Streaming Video Geometry via Dynamic Feature Normalization

Xiaoyang Lyu, Muxin Liu, Xiaoshan Wu +5

Consistent 3D geometry estimation from streaming RGB input is crucial for real-world applications such as autonomous driving, embodied AI, and large-scale reconstruction. While mod…

cs.CV2026

Controllable Video Object Insertion via Multi-View Priors

Qi Xia, Xia Qi, Peishan Cong +4

Video object insertion places a user-specified object in an existing dynamic scene. Existing methods typically condition generation on text or a single reference image. Consequentl…

cs.GR2026

AniGen: Unified Fields for Animatable 3D Asset Generation

Yi-Hua Huang, Zi-Xin Zou, Yuting He +6

Animatable 3D assets, defined as geometry equipped with an articulated skeleton and skinning weights, are fundamental to interactive graphics, embodied agents, and animation produc…