3 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 1 cited
Breathing Life into Faces: Speech-driven 3D Facial Animation with Natural Head Pose and Detailed Shape
Wei Zhao, Yijun Wang, Tianyu He +3
The creation of lifelike speech-driven 3D facial animation requires a natural and precise synchronization between audio input and facial expressions. However, existing works still…
cs.CV2023★ 3 cited
Towards Better Multi-modal Keyphrase Generation via Visual Entity Enhancement and Multi-granularity Image Noise Filtering
Yifan Dong, Suhang Wu, Fandong Meng +4
Multi-modal keyphrase generation aims to produce a set of keyphrases that represent the core points of the input text-image pair. In this regard, dominant methods mainly focus on m…
cs.CV2023★ 3 cited
DiffColor: Toward High Fidelity Text-Guided Image Colorization with Diffusion Models
Jianxin Lin, Peng Xiao, Yijun Wang +2
Recent data-driven image colorization methods have enabled automatic or reference-based colorization, while still suffering from unsatisfactory and inaccurate object-level color co…