4 papers
RapidMV: Leveraging Spatio-Angular Representations for Efficient and Consistent Text-to-Multi-View Synthesis
Seungwook Kim, Yichun Shi, Kejie Li +2
Generating synthetic multi-view images from a text prompt is an essential bridge to generating synthetic 3D assets. In this work, we introduce RapidMV, a novel text-to-multi-view g…
Dual Diffusion for Unified Image Generation and Understanding
Zijie Li, Henry Li, Yichun Shi +4
Diffusion models have gained tremendous success in text-to-image generation, yet still lag behind with visual understanding tasks, an area dominated by autoregressive vision-langua…
MVLight: Relightable Text-to-3D Generation via Light-conditioned Multi-View Diffusion
Dongseok Shim, Yichun Shi, Kejie Li +2
Recent advancements in text-to-3D generation, building on the success of high-performance text-to-image generative models, have made it possible to create imaginative and richly te…
SeedEdit: Align Image Re-Generation to Image Editing
Yichun Shi, Peng Wang, Weilin Huang
We introduce SeedEdit, a diffusion model that is able to revise a given image with any text prompt. In our perspective, the key to such a task is to obtain an optimal balance betwe…