collaborators

7 papers

cs.CV2026

To See a World in a Living Context: Unified Indoor-Outdoor Urban World Generation

Xiaobin Huang, Zilong Huang, Yang Luo +3

Text-driven 3D generation has advanced rapidly in creating large-scale outdoor environments and detailed indoor scenes, but these domains are usually synthesized independently, lac…

cs.CV2026

Local-GS: Accelerating 3D Gaussian Splatting via Tile-Local Warp Coherence

Yang Luo, Yan Gong, Yongsheng Gao +3

3D Gaussian Splatting (3DGS) has significantly advanced real-time novel view synthesis by representing scenes as dense collections of anisotropic 3D Gaussian primitives. However, t…

cs.CV2026

Semantic Prior Guided One-View 6D Pose Estimation for Novel Objects

Yang Luo, Yan Gong, Yongsheng Gao +3

In many practical 6D object pose estimation scenarios, we often have access to only a single real-world RGB-D reference view per object, typically without CAD models. Existing meth…

cs.CV2026

MUSE: Resolving Manifold Misalignment in Visual Tokenization via Topological Orthogonality

Panqi Yang, Haodong Jing, Jiahao Chao +5

Unified visual tokenization faces a fundamental trade-off between high-fidelity pixel reconstruction (spatial equivariance) and semantic abstraction (conceptual invariance). We att…

cs.CV2026

3DCity-LLM: Empowering Multi-modality Large Language Models for 3D City-scale Perception and Understanding

Yiping Chen, Jinpeng Li, Wenyu Ke +6

While multi-modality large language models excel in object-centric or indoor scenarios, scaling them to 3D city-scale environments remains a formidable challenge. To bridge this ga…

cs.CV2025

SGS-3D: High-Fidelity 3D Instance Segmentation via Reliable Semantic Mask Splitting and Growing

Chaolei Wang, Yang Luo, Jing Du +3

Accurate 3D instance segmentation is crucial for high-quality scene understanding in the 3D vision domain. However, 3D instance segmentation based on 2D-to-3D lifting approaches st…