5 citations · 10 across the 3 of their papers we have counts for
6 papers · 1 filter
Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer
Jiahao Cui, Hui Li, Yun Zhan +7
Existing methodologies for animating portrait images face significant challenges, particularly in handling non-frontal perspectives, rendering dynamic objects around the portrait,…
Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
Mingwang Xu, Hui Li, Qingkun Su +6
The field of portrait image animation, driven by speech audio input, has experienced significant advancements in the generation of realistic and dynamic portraits. This research de…
MM-Diff: High-Fidelity Image Personalization via Multi-Modal Condition Integration
Zhichao Wei, Qingkun Su, Long Qin +1
Recent advances in tuning-free personalized image generation based on diffusion models are impressive. However, to improve subject fidelity, existing methods either retrain the dif…
Fine-grained Text-Video Retrieval with Frozen Image Encoders
Zuozhuo Dai, Fangtao Shao, Qingkun Su +2
State-of-the-art text-video retrieval (TVR) methods typically utilize CLIP and cosine similarity for efficient retrieval. Meanwhile, cross attention methods, which employ a transfo…
MeshMVS: Multi-View Stereo Guided Mesh Reconstruction
Rakesh Shrestha, Zhiwen Fan, Qingkun Su +3
Deep learning based 3D shape generation methods generally utilize latent features extracted from color images to encode the semantics of objects and guide the shape generation proc…
Sketch-R2CNN: An Attentive Network for Vector Sketch Recognition
Lei Li, Changqing Zou, Youyi Zheng +3
Freehand sketching is a dynamic process where points are sequentially sampled and grouped as strokes for sketch acquisition on electronic devices. To recognize a sketched object, m…