activity
20242026
collaborators

6 papers

cs.CV2026

Structural Energy Guidance for View-Consistent Text-to-3D Generation

Qing Zhang, Jinguang Tong, Jing Zhang +2

Text-to-3D generation based on diffusion models often suffers from the Janus problem, leading to inconsistent geometry across viewpoints. This work identifies viewpoint bias in 2D…

cs.CV2026

Probing and Bridging Geometry-Interaction Cues for Affordance Reasoning in Vision Foundation Models

Qing Zhang, Xuesong Li, Jing Zhang

What does it mean for a visual system to truly understand affordance? We argue that this understanding hinges on two complementary capacities: geometric perception, which identifie…

cs.CV2026

RnG: A Unified Transformer for Complete 3D Modeling from Partial Observations

Mochu Xiang, Zhelun Shen, Xuesong Li +7

Human perceive the 3D world through 2D observations from limited viewpoints. While recent feed-forward generalizable 3D reconstruction models excel at recovering 3D structures from…

cs.CV2025

Structural Energy-Guided Sampling for View-Consistent Text-to-3D

Qing Zhang, Jinguang Tong, Jie Hong +2

Text-to-3D generation often suffers from the Janus problem, where objects look correct from the front but collapse into duplicated or distorted geometry from other angles. We attri…

cs.CV2025

Improving Viewpoint Consistency in 3D Generation via Structure Feature and CLIP Guidance

Qing Zhang, Jinguang Tong, Jing Zhang +2

Despite recent advances in text-to-3D generation techniques, current methods often suffer from geometric inconsistencies, commonly referred to as the Janus Problem. This paper iden…

cs.HC2024

Quality Control in Open-Ended Crowdsourcing: A Survey

Lei Chai, Hailong Sun, Jing Zhang

Crowdsourcing provides a flexible approach for leveraging human intelligence to solve large-scale problems, gaining widespread acceptance in domains like intelligent information pr…