4 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CV2025
Abstractive Visual Understanding of Multi-modal Structured Knowledge: A New Perspective for MLLM Evaluation
Yichi Zhang, Zhuo Chen, Lingbing Guo +4
Multi-modal large language models (MLLMs) incorporate heterogeneous modalities into LLMs, enabling a comprehensive understanding of diverse scenarios and objects. Despite the proli…
cs.MM2025
Towards Structure-aware Model for Multi-modal Knowledge Graph Completion
Linyu Li, Zhi Jin, Yichi Zhang +5
Knowledge graphs (KGs) play a key role in promoting various multimedia and AI applications. However, with the explosive growth of multi-modal information, traditional knowledge gra…
cs.CV2024★ 4 cited
Unleashing the Potential of SAM2 for Biomedical Images and Videos: A Survey
Yichi Zhang, Zhenrong Shen
The unprecedented developments in segmentation foundational models have become a dominant force in the field of computer vision, introducing a multitude of previously unexplored ca…