3 papers
cs.CV2025
Proxy-Tuning: Tailoring Multimodal Autoregressive Models for Subject-Driven Image Generation
Yi Wu, Shengju Qian, Lingting Zhu +5
Multimodal autoregressive (AR) models, based on next-token prediction and transformer architecture, have demonstrated remarkable capabilities in various multimodal tasks including…
cs.CV2025
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation
Yi Wu, Lingting Zhu, Shengju Qian +4
In the current research landscape, multimodal autoregressive (AR) models have shown exceptional capabilities across various domains, including visual understanding and generation.…
cs.CV2024
One-shot Generative Domain Adaptation in 3D GANs
Ziqiang Li, Yi Wu, Chaoyue Wang +2
3D-aware image generation necessitates extensive training data to ensure stable training and mitigate the risk of overfitting. This paper first considers a novel task known as One-…