collaborators

6 papers

cs.CV2026

P2Voxel: Pyramid Pivot Voxelization for 3D Mesh Tokenization

Zhenhong Sun, Haozhe Liu, Yifu Wang +6

Triangle meshes provide explicit and accurate surface geometry, yet their irregular topology connectivity makes 3D mesh tokenization a geometric sampling problem: how to sample and…

cs.GR2026

Hi-TOPS: Hierarchical Topology-aware Scoring Prior for 3D Part Decomposition

Ruoyu Wu, Zhenhong Sun, Xiaoming Gong +5

Accurate 3D part decomposition requires separating shapes into structurally meaningful components with precise boundaries while preserving articulation seams and thin attachments.…

cs.AI2026

Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds

Jiaming Bian, Bingliang Li, Yuehao Wu +5

As embodied AI and world models increasingly operate in dynamic 3D environments, visual perception must move beyond passively interpreting given observations toward actively decidi…

cs.CV2026

Social Structure Matters in 3D Human-Human Interaction Generation

Zhongju Wang, Beier Wang, Yatao Bian +6

Although text-to-motion generation has achieved strong progress in synthesizing realistic single-person motions from language, extending it to text-driven 3D human-human interactio…

cs.CV2026

StoryBlender: Inter-Shot Consistent and Editable 3D Storyboard with Spatial-temporal Dynamics

Bingliang Li, Zhenhong Sun, Jiaming Bian +6

Storyboarding is a core skill in visual storytelling for film, animation, and games. However, automating this process requires a system to achieve two properties that current appro…

cs.CV2025

Hierarchical and Step-Layer-Wise Tuning of Attention Specialty for Multi-Instance Synthesis in Diffusion Transformers

Chunyang Zhang, Zhenhong Sun, Zhicheng Zhang +5

Text-to-image (T2I) generation models often struggle with multi-instance synthesis (MIS), where they must accurately depict multiple distinct instances in a single image based on c…