3 papers
cs.CV2024
CMAL: A Novel Cross-Modal Associative Learning Framework for Vision-Language Pre-Training
Zhiyuan Ma, Jianjun Li, Guohui Li +1
With the flourishing of social media platforms, vision-language pre-training (VLP) recently has received great attention and many remarkable progresses have been achieved. The succ…
cs.IR2024
DualVAE: Dual Disentangled Variational AutoEncoder for Recommendation
Zhiqiang Guo, Guohui Li, Jianjun Li +2
Learning precise representations of users and items to fit observed interaction data is the fundamental task of collaborative filtering. Existing studies usually infer entangled re…
cs.IR2023
LGMRec: Local and Global Graph Learning for Multimodal Recommendation
Zhiqiang Guo, Jianjun Li, Guohui Li +3
The multimodal recommendation has gradually become the infrastructure of online media platforms, enabling them to provide personalized service to users through a joint modeling of…