5 citations · 6 across the 6 of their papers we have counts for
6 papers
Towards Automatic Soccer Commentary Generation with Knowledge-Enhanced Visual Reasoning
Zeyu Jin, Xiaoyu Qin, Songtao Zhou +2
Soccer commentary plays a crucial role in enhancing the soccer game viewing experience for audiences. Previous studies in automatic soccer commentary generation typically adopt an…
From Natural Alignment to Conditional Controllability in Multimodal Dialogue
Zeyu Jin, Songtao Zhou, Haoyu Wang +5
The recent advancement of Artificial Intelligence Generated Content (AIGC) has led to significant strides in modeling human interaction, particularly in the context of multimodal d…
V-CASS: Vision-context-aware Expressive Speech Synthesis for Enhancing User Understanding of Videos
Qixin Wang, Songtao Zhou, Zeyu Jin +3
Automatic video commentary systems are widely used on multimedia social media platforms to extract factual information about video content. However, current systems may overlook es…
DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera Synthesis
Zixuan Wang, Jiayi Li, Xiaoyu Qin +4
Synthesizing camera movements from music and dance is highly challenging due to the contradicting requirements and complexities of dance cinematography. Unlike human movements, whi…
Speech-Driven 3D Face Animation with Composite and Regional Facial Movements
Haozhe Wu, Songtao Zhou, Jia Jia +3
Speech-driven 3D face animation poses significant challenges due to the intricacy and variability inherent in human facial movements. This paper emphasizes the importance of consid…
Open-Access Data and Toolbox for Tracking COVID-19 Impact on Power Systems
Guangchun Ruan, Zekuan Yu, Shutong Pu +5
Intervention policies against COVID-19 have caused large-scale disruptions globally, and led to a series of pattern changes in the power system operation. Analyzing these pandemic-…