2 papers
cs.IR2024
QARM: Quantitative Alignment Multi-Modal Recommendation at Kuaishou
Xinchen Luo, Jiangxia Cao, Tianyu Sun +17
In recent years, with the significant evolution of multi-modal large models, many recommender researchers realized the potential of multi-modal information for user interest modeli…
cs.CV2023
Generation-Guided Multi-Level Unified Network for Video Grounding
Xing Cheng, Xiangyu Wu, Dong Shen +2
Video grounding aims to locate the timestamps best matching the query description within an untrimmed video. Prevalent methods can be divided into moment-level and clip-level frame…