activity
20242026
collaborators

31 papers

cs.GR2026

MultiCube: Compositional 3D Generation With Part-Level Semantic and Spatial Control

Ava Pun, Kangle Deng, Yiheng Zhu +3

Digital 3D objects used in games and animation are often required to be compositional; that is, decomposed into semantically meaningful parts. Recent 3D generation methods can prod…

cs.HC2026

Collascope: Supporting Serendipitous Asset Exploration for Collage-Based Storytelling

Jiayi Zhou, Longji Huang, Lvmin Zhang +5

Collage-based storytelling requires visual elements that support emerging narratives and inspire creative reinterpretation. Existing tools, however, rely largely on keyword- and im…

cs.CV2026

Masked Visual Actions for Unified World Modeling

Hadi Alzayer, Wenlong Huang, Haonan Chen +8

Video models absorb rich priors over how the visual world moves, interacts, and responds to contact, making them promising substrates for robotic world modeling. The central challe…

cs.LG2026

FourTune: Towards Fully 4-Bit Efficient Post-Training for Diffusion Models

Bowen Xue, Zihan Min, Xingyang Li +8

Diffusion models have become a dominant paradigm for high-quality generative modeling, while post-training is essential for adapting them to diverse downstream applications. Howeve…

cs.CV2026

TinyHistory: Lightweight Video History Embeddings via Two-Stage Context Learning

Lvmin Zhang, Shengqu Cai, Muyang Li +6

History context is central to autoregressive video generation, driving consistency and storytelling for both commercial models and personal use cases. For example, personal users,…

cs.GR2026

Aggregating LLM-Based Weak Verifiers for Spatial Layout Generation

Sharon Zhang, R. Kenny Jones, Jiajun Wu +1

We present a pipeline for building and aggregating task-specific, LLM-generated weak (imperfect) verifiers into a strong verifier for spatial layout domains. Given a task descripti…