MSCRS: Multi-modal Semantic Graph Prompt Learning Framework for Conversational Recommender Systems
arXiv:2504.10921 · doi:10.1145/3726302.3730040
Abstract
Conversational Recommender Systems (CRSs) aim to provide personalized recommendations by interacting with users through conversations. Most existing studies of CRS focus on extracting user preferences from conversational contexts. However, due to the short and sparse nature of conversational contexts, it is difficult to fully capture user preferences by conversational contexts only. We argue that multi-modal semantic information can enrich user preference expressions from diverse dimensions (e.g., a user preference for a certain movie may stem from its magnificent visual effects and compelling storyline). In this paper, we propose a multi-modal semantic graph prompt learning framework for CRS, named MSCRS. First, we extract textual and image features of items mentioned in the conversational contexts. Second, we capture higher-order semantic associations within different semantic modalities (collaborative, textual, and image) by constructing modality-specific graph structures. Finally, we propose an innovative integration of multi-modal semantic graphs with prompt learning, harnessing the power of large language models to comprehensively explore high-dimensional semantic relationships. Experimental results demonstrate that our proposed method significantly improves accuracy in item recommendation, as well as generates more natural and contextually relevant content in response generation.
References in corpus (12)
- Mining Latent Structures for Multimedia Recommendation
- Bootstrap Latent Representations for Multi-modal Recommendation
- Estimation-Action-Reflection: Towards Deep Interaction Between Conversational and Recommender Systems
- Multi-Modal Self-Supervised Learning for Recommendation
- Interactive Path Reasoning on Graph for Conversational Recommendation
- Towards Unified Conversational Recommender Systems via Knowledge-Enhanced Prompt Learning
- Seamlessly Unifying Attributes and Items: Conversational Recommendation for Cold-Start Users
- Towards Question-based Recommender Systems
- User-Centric Conversational Recommendation with Multi-Aspect User Modeling
- Variational Reasoning over Incomplete Knowledge Graphs for Conversational Recommendation
- Knowledge Graphs and Pre-trained Language Models enhanced Representation Learning for Conversational Recommender Systems
- Learning to Ask: Conversational Product Search via Representation Learning