Research
Pre-trained LLMs Meet Sequential Recommenders: Efficient User-Centric Knowledge Distillation
arXiv:2604.21536v1 Announce Type: cross Abstract: Sequential recommender systems have achieved significant success in modeling temporal user behavior but remain limited in capturing rich user semantic
arXiv:2604.21536v1 Announce Type: cross Abstract: Sequential recommender systems have achieved significant success in modeling temporal user behavior but remain limited in capturing rich user semantics beyond interaction patterns. Large Language Models (LLMs) present opportunities to enhance user understanding with their reasoning capabilities, yet existing integration approaches create prohibitive inference costs in real time. To address these limitations, we present a novel knowledge distillation method that utilizes textual user profile generated by pre-trained LLMs into sequential recommenders without requiring LLM inference at serving time. The resulting approach maintains the inference efficiency of traditional sequential models while requiring neither architectural modifications nor LLM fine-tuning.
Related
- TokenFormer: Unify the Multi-Field and Sequential Recommendation Worlds
- ItemRAG: Item-Based Retrieval-Augmented Generation for LLM-Based Recommendation
- Distilling Genomic Models for Efficient mRNA Representation Learning via Embedding Matching
Source: arXiv cs.AI | 2026-04-24