activity
20242026
collaborators

12 papers

cs.CV2026

Interacted Planes Reveal 3D Line Mapping

Zeran Ke, Bin Tan, Gui-Song Xia +2

3D line mapping from multi-view RGB images provides a compact and structured visual representation of scenes. We study the problem from a physical and topological perspective: a 3D…

cs.CV2026

Masked Depth Modeling for Spatial Perception

Bin Tan, Changjiang Sun, Xiage Qin +8

Spatial visual perception is a fundamental requirement in physical-world applications like autonomous driving and robotic manipulation, driven by the need to interact with 3D envir…

cs.CV2026

FlowSSC: Universal Generative Monocular Semantic Scene Completion via One-Step Latent Diffusion

Zichen Xi, Hao-Xiang Chen, Nan Xue +5

Semantic Scene Completion (SSC) from monocular RGB images is a fundamental yet challenging task due to the inherent ambiguity of inferring occluded 3D geometry from a single view.…

cs.CV2025

Real-Time 3D Object Detection with Inference-Aligned Learning

Chenyu Zhao, Xianwei Zheng, Zimin Xia +2

Real-time 3D object detection from point clouds is essential for dynamic scene understanding in applications such as augmented reality, robotics and navigation. We introduce a nove…

cs.CV2025

PLANA3R: Zero-shot Metric Planar 3D Reconstruction via Feed-Forward Planar Splatting

Changkun Liu, Bin Tan, Zeran Ke +6

This paper addresses metric 3D reconstruction of indoor scenes by exploiting their inherent geometric regularities with compact representations. Using planar 3D primitives - a well…

cs.CV2025

UCD: Unconditional Discriminator Promotes Nash Equilibrium in GANs

Mengfei Xia, Nan Xue, Jiapeng Zhu +1

Adversarial training turns out to be the key to one-step generation, especially for Generative Adversarial Network (GAN) and diffusion model distillation. Yet in practice, GAN trai…