2 papers
cs.AI2026
MGDT: MLLM-Guided Diffusion Transformer with Relation-Adaptive Mixture-of-Experts for Multimodal Knowledge Graph Completion
Xu Hou, Meiyu Liang, Wei Huang +6
Multimodal Knowledge Graph Completion (MKGC) requires inferring missing entities from structural, textual, and visual cues. Existing diffusion-based MKGC methods usually denoise di…
cs.CV2025
LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representation
Xin Lu, Chuanqing Zhuang, Chenxi Jin +4
Speech-driven 3D facial animation has attracted increasing interest since its potential to generate expressive and temporally synchronized digital humans. While recent works have b…