3 papers
cs.CL2026
AdaMTP: An Adaptive Training Paradigm for Multi-Token Prediction
Ziqiang Cui, Han Shi, Bowei He +8
Multi-Token Prediction (MTP) has emerged as an effective paradigm that augments a shared Large Language Model backbone with auxiliary heads, training the model to predict several f…
cs.IR2024
Fusion Matters: Learning Fusion in Deep Click-through Rate Prediction Models
Kexin Zhang, Fuyuan Lyu, Xing Tang +5
The evolution of previous Click-Through Rate (CTR) models has mainly been driven by proposing complex components, whether shallow or deep, that are adept at modeling feature intera…
cs.IR2024
Comprehending Knowledge Graphs with Large Language Models for Recommender Systems
Ziqiang Cui, Yunpeng Weng, Xing Tang +4
In recent years, the introduction of knowledge graphs (KGs) has significantly advanced recommender systems by facilitating the discovery of potential associations between items. Ho…