4 papers
Modularized Dynamic-Granularity Video LLM for Multi-Event Long Video Understanding
Wei Feng, Xin Wang, Yu-Wei Zhan +2
Video Large Language Models (Video LLMs) have made significant advancements in various video understanding tasks. However, long-video scenarios remain challenging due to the tensio…
Out-of-Distribution Generalization in Graph Foundation Models
Haoyang Li, Haibo Chen, Xin Wang +1
Graphs are a fundamental data structure for representing relational information in domains such as social networks, molecular systems, and knowledge graphs. However, graph learning…
Towards Multimodal Graph Large Language Model
Xin Wang, Zeyang Zhang, Linxin Xiao +3
Multi-modal graphs, which integrate diverse multi-modal features and relations, are ubiquitous in real-world applications. However, existing multi-modal graph learning methods are…
Modular Machine Learning: An Indispensable Path towards New-Generation Large Language Models
Xin Wang, Haoyang Li, Haibo Chen +2
Large language models (LLMs) have substantially advanced machine learning research, including natural language processing, computer vision, data mining, etc., yet they still exhibi…