Aug 2026
RosePO: Customized Preference Alignment in LLM-Based Recommendation
This work proposes RosePO, a framework to refine LLM-based recommendation through pairwise preference optimization with personalized smoothing, and incorporates a personalized smoothing factor predicted by a user oracle into the optimization objective.
Jiayi Liao, Xiangnan He, Ruobing Xie et al.
· ACM Transactions on Informat... · 0 citations