6 papers
CamPVG: Camera-Controlled Panoramic Video Generation with Epipolar-Aware Diffusion
Chenhao Ji, Chaohui Yu, Junyao Gao +2
Recently, camera-controlled video generation has seen rapid development, offering more precise control over video generation. However, existing methods predominantly focus on camer…
CharacterShot: Controllable and Consistent 4D Character Animation
Junyao Gao, Jiaxing Li, Wenran Liu +5
In this paper, we propose \textbf{CharacterShot}, a controllable and consistent 4D character animation framework that enables any individual designer to create dynamic 3D character…
One Object, Multiple Lies: A Benchmark for Cross-task Adversarial Attack on Unified Vision-Language Models
Jiale Zhao, Xinyang Jiang, Junyao Gao +2
Unified vision-language models(VLMs) have recently shown remarkable progress, enabling a single model to flexibly address diverse tasks through different instructions within a shar…
TRAIL: Transferable Robust Adversarial Images via Latent diffusion
Yuhao Xue, Zhifei Zhang, Xinyang Jiang +6
Adversarial attacks exploiting unrestricted natural perturbations present severe security risks to deep learning systems, yet their transferability across models remains limited du…
FaceShot: Bring Any Character into Life
Junyao Gao, Yanan Sun, Fei Shen +4
In this paper, we present FaceShot, a novel training-free portrait animation framework designed to bring any character into life from any driven video without fine-tuning or retrai…
DiffPano: Scalable and Consistent Text to Panorama Generation with Spherical Epipolar-Aware Diffusion
Weicai Ye, Chenhao Ji, Zheng Chen +7
Diffusion-based methods have achieved remarkable achievements in 2D image or 3D object generation, however, the generation of 3D scenes and even images remains constr…