12 citations · 28 across the 7 of their papers we have counts for
11 papers
Towards Secure and Usable 3D Assets: A Novel Framework for Automatic Visible Watermarking
Gursimran Singh, Tianxi Hu, Mohammad Akbari +2
3D models, particularly AI-generated ones, have witnessed a recent surge across various industries such as entertainment. Hence, there is an alarming need to protect the intellectu…
StereoCrafter: Diffusion-based Generation of Long and High-fidelity Stereoscopic 3D from Monocular Videos
Sijie Zhao, Wenbo Hu, Xiaodong Cun +6
This paper presents a novel framework for converting 2D videos to immersive stereoscopic 3D, addressing the growing demand for 3D content in immersive experience. Leveraging founda…
Fast Gradient Computation for Gromov-Wasserstein Distance
Wei Zhang, Zihao Wang, Jie Fan +2
The Gromov-Wasserstein distance is a notable extension of optimal transport. In contrast to the classic Wasserstein distance, it solves a quadratic assignment problem that minimize…
Boosting Chinese ASR Error Correction with Dynamic Error Scaling Mechanism
Jiaxin Fan, Yong Zhang, Hanzhang Li +5
Chinese Automatic Speech Recognition (ASR) error correction presents significant challenges due to the Chinese language's unique features, including a large character set and borde…
Make-Your-Video: Customized Video Generation Using Textual and Structural Guidance
Jinbo Xing, Menghan Xia, Yuxin Liu +9
Creating a vivid video from the event or scenario in our imagination is a truly fascinating experience. Recent advancements in text-to-video synthesis have unveiled the potential t…
Inserting Anybody in Diffusion Models via Celeb Basis
Ge Yuan, Xiaodong Cun, Yong Zhang +5
Exquisite demand exists for customizing the pretrained large text-to-image model, , Stable Diffusion, to generate innovative concepts, such as the users themselves.…