Pre-trained LLMs Meet Sequential Recommenders: Efficient User-Centric Knowledge Distillation
arXiv:2604.21536 · doi:10.1007/978-3-032-21300-6_42
Abstract
Sequential recommender systems have achieved significant success in modeling temporal user behavior but remain limited in capturing rich user semantics beyond interaction patterns. Large Language Models (LLMs) present opportunities to enhance user understanding with their reasoning capabilities, yet existing integration approaches create prohibitive inference costs in real time. To address these limitations, we present a novel knowledge distillation method that utilizes textual user profile generated by pre-trained LLMs into sequential recommenders without requiring LLM inference at serving time. The resulting approach maintains the inference efficiency of traditional sequential models while requiring neither architectural modifications nor LLM fine-tuning.
Accepted to ECIR 2026. 7 pages. This version of the contribution has been accepted for publication, after peer review but is not the Version of Record and does not reflect post-acceptance improvements, or any corrections. The Version of Record is available online at: http://dx.doi.org/10.1007/978-3-032-21300-6_42
References in corpus (6)
- Predicting Dynamic Embedding Trajectory in Temporal Interaction Networks
- A Critical Study on Data Leakage in Recommender System Offline Evaluation
- Turning Dross Into Gold Loss: is BERT4Rec really better than SASRec?
- Does It Look Sequential? An Analysis of Datasets for Evaluation of Sequential Recommendations
- Time to Split: Exploring Data Splitting Strategies for Offline Evaluation of Sequential Recommenders
- eSASRec: Enhancing Transformer-based Recommendations in a Modular Fashion