works on

From the 2 of 6 linked papers with an AI index.

collaborators

6 papers

cs.CV2026

From Hindsight to Foresight: Self-Encouraged Hindsight Distillation for Knowledge-based Visual Question Answering

Yu Zhao, Ying Zhang, Xuhui Sui +4

The paper introduces a teacher‑student framework called Hindsight Distillation (HinD) that uses privileged answer information to generate reasoning trajectories for a multimodal LL…

cs.LG2026

Hyper-modal Imputation Diffusion Embedding with Dual-Distillation for Federated Multimodal Knowledge Graph Completion

Ying Zhang, Yu Zhao, Xuhui Sui +5

The paper introduces a federated learning framework for multimodal knowledge graph completion that recovers missing multimodal information with a hyper-modal imputation diffusion e…

cs.MM2026

Hyperbolic Multimodal Generative Representation Learning for Generalized Zero-Shot Multimodal Information Extraction

Baohang Zhou, Kehui Song, Rize Jin +5

Multimodal information extraction (MIE) constitutes a set of essential tasks aimed at extracting structural information from Web texts with integrating images, to facilitate the st…

cs.CL2025

Plan of Knowledge: Retrieval-Augmented Large Language Models for Temporal Knowledge Graph Question Answering

Xinying Qian, Ying Zhang, Yu Zhao +3

Temporal Knowledge Graph Question Answering (TKGQA) aims to answer time-sensitive questions by leveraging factual information from Temporal Knowledge Graphs (TKGs). While previous…

cs.MM2025

Dark Side of Modalities: Reinforced Multimodal Distillation for Multimodal Knowledge Graph Reasoning

Yu Zhao, Ying Zhang, Xuhui Sui +4

The multimodal knowledge graph reasoning (MKGR) task aims to predict the missing facts in the incomplete MKGs by leveraging auxiliary images and descriptions of entities. Existing…

cs.MM2025

Multimodal Graph-Based Variational Mixture of Experts Network for Zero-Shot Multimodal Information Extraction

Baohang Zhou, Ying Zhang, Yu Zhao +2

Multimodal information extraction on social media is a series of fundamental tasks to construct the multimodal knowledge graph. The tasks aim to extract the structural information…