3 papers
cs.AI2026
MGDT: MLLM-Guided Diffusion Transformer with Relation-Adaptive Mixture-of-Experts for Multimodal Knowledge Graph Completion
Xu Hou, Meiyu Liang, Wei Huang +6
Multimodal Knowledge Graph Completion (MKGC) requires inferring missing entities from structural, textual, and visual cues. Existing diffusion-based MKGC methods usually denoise di…
cs.CV2025
LSF-Animation: Label-Free Speech-Driven Facial Animation via Implicit Feature Representation
Xin Lu, Chuanqing Zhuang, Chenxi Jin +4
Speech-driven 3D facial animation has attracted increasing interest since its potential to generate expressive and temporally synchronized digital humans. While recent works have b…
cs.CV2024
It Takes Two: Accurate Gait Recognition in the Wild via Cross-granularity Alignment
Jinkai Zheng, Xinchen Liu, Boyue Zhang +4
Existing studies for gait recognition primarily utilized sequences of either binary silhouette or human parsing to encode the shapes and dynamics of persons during walking. Silhoue…