collaborators

5 papers

cs.CV2025

Efficient Depth-Guided Urban View Synthesis

Sheng Miao, Jiaxin Huang, Dongfeng Bai +4

Recent advances in implicit scene representation enable high-fidelity street view novel view synthesis. However, existing methods optimize a neural radiance field for each scene, r…

cs.CV2024

An Efficient Occupancy World Model via Decoupled Dynamic Flow and Image-assisted Training

Haiming Zhang, Ying Xue, Xu Yan +6

The field of autonomous driving is experiencing a surge of interest in world models, which aim to predict potential future scenarios based on historical observations. In this paper…

cs.CV2024

DisEnvisioner: Disentangled and Enriched Visual Prompt for Customized Image Generation

Jing He, Haodong Li, Yongzhe Hu +4

In the realm of image generation, creating customized images from visual prompt with additional textual instruction emerges as a promising endeavor. However, existing methods, both…

cs.CV2024

OmniBooth: Learning Latent Control for Image Synthesis with Multi-modal Instruction

Leheng Li, Weichao Qiu, Xu Yan +6

We present OmniBooth, an image generation framework that enables spatial control with instance-level multi-modal customization. For all instances, the multimodal instruction can be…

cs.CV2024

SyntheOcc: Synthesize Geometric-Controlled Street View Images through 3D Semantic MPIs

Leheng Li, Weichao Qiu, Yingjie Cai +4

The advancement of autonomous driving is increasingly reliant on high-quality annotated datasets, especially in the task of 3D occupancy prediction, where the occupancy labels requ…