Progressive Alignment of Recommender Foundation Model through Multi-Phase Post-Training
通过多阶段后训练实现推荐系统基础模型的渐进式对齐
专题命中 偏好对齐 :alignment(title,abstract);分类 cs.AI
AI总结 本文提出LP-FFT-RFT三阶段渐进式后训练框架,将推荐系统基础模型的下游适配与业务指标对齐分离,实验证实其能提升推荐质量。
Comments 9 pages, 3 figures. Accepted to the 20th ACM Conference on Recommender Systems (RecSys '26), Industry Track