5 papers
HyperSketch: Controllable Video Sketching in a Style Hyperspace
Xinding Zhu, Xinye Yang, Yingping Yang +3
Vector sketch animation offers tremendous advantages for multimedia and creative design through concise line expressions and flexible editing. Learning-based generation methods of…
CaLa: Complementary Association Learning for Augmenting Composed Image Retrieval
Xintong Jiang, Yaxiong Wang, Mengjian Li +3
Composed Image Retrieval (CIR) involves searching for target images based on an image-text pair query. While current methods treat this as a query-target matching problem, we argue…
Speech-driven Personalized Gesture Synthetics: Harnessing Automatic Fuzzy Feature Inference
Fan Zhang, Zhaohan Wang, Xin Lyu +8
Speech-driven gesture generation is an emerging field within virtual human creation. However, a significant challenge lies in accurately determining and processing the multitude of…
Towards Geometric-Photometric Joint Alignment for Facial Mesh Registration
Xizhi Wang, Yaxiong Wang, Mengjian Li
This paper presents a Geometric-Photometric Joint Alignment~(GPJA) method, which aligns discrete human expressions at pixel-level accuracy by combining geometric and photometric in…
POS: A Prompts Optimization Suite for Augmenting Text-to-Video Generation
Shijie Ma, Huayi Xu, Mengjian Li +3
This paper targets to enhance the diffusion-based text-to-video generation by improving the two input prompts, including the noise and the text. Accommodated with this goal, we pro…