1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.LG2025
Decoupling and Damping: Structurally-Regularized Gradient Matching for Multimodal Graph Condensation
Lian Shen, Zhendan Chen, Meijia Song +4
In multimodal graph learning, graph structures that integrate information from multiple sources, such as vision and text, can more comprehensively model complex entity relationship…
cs.CV2025★ 1 cited
Fine-Tuning Vision-Language Models for Visual Navigation Assistance
Xiao Li, Bharat Gandhi, Ming Zhan +6
We address vision-language-driven indoor navigation to assist visually impaired individuals in reaching a target location using images and natural language guidance. Traditional na…