most citedInserting Anybody in Diffusion Models via Celeb Basis

12 citations · 28 across the 7 of their papers we have counts for

collaborators

11 papers

cs.CV2024

Towards Secure and Usable 3D Assets: A Novel Framework for Automatic Visible Watermarking

Gursimran Singh, Tianxi Hu, Mohammad Akbari +2

3D models, particularly AI-generated ones, have witnessed a recent surge across various industries such as entertainment. Hence, there is an alarming need to protect the intellectu…

cs.CV20242 cited

StereoCrafter: Diffusion-based Generation of Long and High-fidelity Stereoscopic 3D from Monocular Videos

Sijie Zhao, Wenbo Hu, Xiaodong Cun +6

This paper presents a novel framework for converting 2D videos to immersive stereoscopic 3D, addressing the growing demand for 3D content in immersive experience. Leveraging founda…

cs.LG2024

Fast Gradient Computation for Gromov-Wasserstein Distance

Wei Zhang, Zihao Wang, Jie Fan +2

The Gromov-Wasserstein distance is a notable extension of optimal transport. In contrast to the classic Wasserstein distance, it solves a quadratic assignment problem that minimize…

cs.CL20231 cited

Boosting Chinese ASR Error Correction with Dynamic Error Scaling Mechanism

Jiaxin Fan, Yong Zhang, Hanzhang Li +5

Chinese Automatic Speech Recognition (ASR) error correction presents significant challenges due to the Chinese language's unique features, including a large character set and borde…

cs.CV20232 cited

Make-Your-Video: Customized Video Generation Using Textual and Structural Guidance

Jinbo Xing, Menghan Xia, Yuxin Liu +9

Creating a vivid video from the event or scenario in our imagination is a truly fascinating experience. Recent advancements in text-to-video synthesis have unveiled the potential t…

cs.CV202312 cited

Inserting Anybody in Diffusion Models via Celeb Basis

Ge Yuan, Xiaodong Cun, Yong Zhang +5

Exquisite demand exists for customizing the pretrained large text-to-image model, , Stable Diffusion, to generate innovative concepts, such as the users themselves.…