4 papers
Structural Energy Guidance for View-Consistent Text-to-3D Generation
Qing Zhang, Jinguang Tong, Jing Zhang +2
Text-to-3D generation based on diffusion models often suffers from the Janus problem, leading to inconsistent geometry across viewpoints. This work identifies viewpoint bias in 2D…
Probing and Bridging Geometry-Interaction Cues for Affordance Reasoning in Vision Foundation Models
Qing Zhang, Xuesong Li, Jing Zhang
What does it mean for a visual system to truly understand affordance? We argue that this understanding hinges on two complementary capacities: geometric perception, which identifie…
Structural Energy-Guided Sampling for View-Consistent Text-to-3D
Qing Zhang, Jinguang Tong, Jie Hong +2
Text-to-3D generation often suffers from the Janus problem, where objects look correct from the front but collapse into duplicated or distorted geometry from other angles. We attri…
Improving Viewpoint Consistency in 3D Generation via Structure Feature and CLIP Guidance
Qing Zhang, Jinguang Tong, Jing Zhang +2
Despite recent advances in text-to-3D generation techniques, current methods often suffer from geometric inconsistencies, commonly referred to as the Janus Problem. This paper iden…