45 citations · 47 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 45 cited
Plug-and-Play Regulators for Image-Text Matching
Haiwen Diao, Ying Zhang, Wei Liu +2
Exploiting fine-grained correspondence and visual-semantic alignments has shown great potential in image-text matching. Generally, recent approaches first employ a cross-modal atte…
cs.CV2023
Audio2Gestures: Generating Diverse Gestures from Audio
Jing Li, Di Kang, Wenjie Pei +4
People may perform diverse gestures affected by various mental and physical factors when speaking the same sentences. This inherent one-to-many relationship makes co-speech gesture…
cs.AI2022★ 2 cited
Enabling Harmonious Human-Machine Interaction with Visual-Context Augmented Dialogue System: A Review
Hao Wang, Bin Guo, Yating Zeng +5
The intelligent dialogue system, aiming at communicating with humans harmoniously with natural language, is brilliant for promoting the advancement of human-machine interaction in…