2 papers
cs.CV2026
HyLoVQA: Dynamic Hypernetwork-Generated Low-Rank Adaptation for Continual Visual Question Answering
Yiran Wang, Chenyi Xiong, Ziyue Qin +3
Continual Visual Question Answering (VQA) requires learning from non-stationary streams of visual inputs and questions while preserving past knowledge. Most prior methods adapt by…
cs.AI2026
MyGram: Modality-aware Graph Transformer with Global Distribution for Multi-modal Entity Alignment
Zhifei Li, Ziyue Qin, Xiangyu Luo +6
Multi-modal entity alignment aims to identify equivalent entities between two multi-modal Knowledge graphs by integrating multi-modal data, such as images and text, to enrich the s…