4 papers · 1 filter
Attribute Token Arithmetic: Disentangled and Continuous Semantic Control for Visual Autoregressive Models
Xindi Yang, Yicheng Wu, Cheng Zhang +2
Autoregressive text-to-image generation has recently achieved remarkable progress, offering high-fidelity synthesis via a unified generative framework. However, fine-grained semant…
VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior
Xindi Yang, Baolu Li, Yiming Zhang +8
Video diffusion models (VDMs) have advanced significantly in recent years, enabling the generation of highly realistic videos and drawing the attention of the community in their po…
Neural Field Classifiers via Target Encoding and Classification Loss
Xindi Yang, Zeke Xie, Xiong Zhou +6
Neural field methods have seen great progress in various long-standing tasks in computer vision and computer graphics, including novel view synthesis and geometry reconstruction. A…
S3IM: Stochastic Structural SIMilarity and Its Unreasonable Effectiveness for Neural Fields
Zeke Xie, Xindi Yang, Yujie Yang +5
Recently, Neural Radiance Field (NeRF) has shown great success in rendering novel-view images of a given scene by learning an implicit representation with only posed RGB images. Ne…