3 papers
cs.CV2026
SemanticNVS: Improving Semantic Scene Understanding in Generative Novel View Synthesis
Xinya Chen, Christopher Wewer, Jiahao Xie +2
We present SemanticNVS, a camera-conditioned multi-view diffusion model for novel view synthesis (NVS), which improves generation quality and consistency by integrating pre-trained…
cs.CV2026
ScenDi: 3D-to-2D Scene Diffusion Cascades for Urban Generation
Hanlei Guo, Jiahao Shao, Xinya Chen +4
Recent advancements in 3D object generation using diffusion models have achieved remarkable success, but generating realistic 3D urban scenes remains challenging. Existing methods…
cs.CV2025
NormalCrafter: Learning Temporally Consistent Normals from Video Diffusion Priors
Yanrui Bin, Wenbo Hu, Haoyuan Wang +2
Surface normal estimation serves as a cornerstone for a spectrum of computer vision applications. While numerous efforts have been devoted to static image scenarios, ensuring tempo…