GOD: Enhancing Generalization via Deep Grafting for Sequential Recommendation

August 17, 2026 ยท Grace Period ยท ๐Ÿ› CIKM 2026 full research paper

โณ Grace Period
This paper is less than 90 days old. We give authors time to release their code before passing judgment.
Authors WooJoo Kim, JunYoung Kim, JaeHyung Lim, HwanJo Yu arXiv ID 2608.16073 Category cs.IR: Information Retrieval Cross-listed cs.LG Citations 0 Venue CIKM 2026 full research paper
Abstract
Sequential recommenders often struggle with sparse and noisy histories, limiting generalization to unseen interactions. Knowledge distillation mitigates this by transferring dense supervision from a teacher to a student. However, most distillation methods run teacher and student independently, then match student outputs or representations to the teacher. Such supervision entangles student-component effects, blurring whether weak generalization stems from unreliable embeddings, overfitted encoding, or co-adaptation to sparse histories. In this paper, we propose Graft-Oriented Distillation (GOD), a component-level distillation framework for improved generalization through grafting. Grafting denotes replacing selected frozen-teacher components with trainable student counterparts to build hybrid source models. GOD uses these hybrid models to evaluate student embeddings with the teacher encoder and the student encoder with teacher embeddings, providing component-level feedback. At inference, GOD uses only the student, incurring no additional cost. Across three real-world datasets, GOD outperforms state-of-the-art baselines by up to 13.92%.
Community shame:
Not yet rated
Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

๐Ÿ“œ Similar Papers

In the same crypt โ€” Information Retrieval