2 papers
cs.CV2025
RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation
Xianfeng Tan, Yuhan Li, Wenxiang Shang +6
Standard clothing asset generation involves restoring forward-facing flat-lay garment images displayed on a clear background by extracting clothing information from diverse real-wo…
cs.CV2025
SyncAnimation: A Real-Time End-to-End Framework for Audio-Driven Human Pose and Talking Head Animation
Yujian Liu, Shidang Xu, Jing Guo +4
Generating talking avatar driven by audio remains a significant challenge. Existing methods typically require high computational costs and often lack sufficient facial detail and r…