Skip to content

Multi-Decoder OneRec: Controllable Generative Retrieval for Multi-Objective Industrial Recommendation

Jul 2026 · arXiv.org · Vol abs/2607.26500 · 0 citations · 35 references
Computer Science

TL;DR

Results show that generative retrieval can combine shared modeling with objective-specific control and complementary candidate generation, and under the same 512-item retrieval budget, Multi-Decoder OneRec improves over the single-decoder OneRec baseline.

Abstract

Industrial recommender systems build candidate pools by assigning explicit quotas to objective-specific retrieval routes. This design offers quota control but increasingly fragments modeling, training, and serving as the route set grows. Semantic-ID-based generative retrieval provides a unified alternative, yet a single decoder entangles objective policies and limits candidate complementarity. We propose Multi-Decoder OneRec, a controllable framework that combines shared representations, isolated objective adaptation, and coordinated decoding. All objectives share a user-context module and the General Decoder, while each objective adds an isolated, parameter-efficient LoRA expert. During training, exposure-sample next-token prediction (NTP) updates the shared base, target-filtered NTP updates the event-based experts, and Kullback-Leibler (KL)-regularized policy optimization updates the Watch-time expert; gradient routing isolates these updates, and the General Decoder supplies a stop-gradient reference. At inference, explicit route quotas allocate the fixed budget and Multi-Decoder Constrained Beam Search reduces cross-route overlap. We publicly release Kwai26, a large-scale multi-objective benchmark with 1.31 billion raw item-level records, 31.85 million Item-ID entries, and 25.03 million items with valid Semantic IDs, together with predefined splits and an evaluation protocol. Under the same 512-item retrieval budget, Multi-Decoder OneRec improves over the single-decoder OneRec baseline by 1.69%-5.62% across four Recall@512 metrics. In a production A/B test, it yields relative gains of 0.37% in app usage time per device, 0.19% in Day-7 retained users, 0.19% in devices with at least one share, and 2.09% in new-content Cold-Start. These results show that generative retrieval can combine shared modeling with objective-specific control and complementary candidate generation.

View source

Similar papers

Preprint Sep 2026

GRP v0.1 Technical Report

Industrial recommendation systems rely on multi-stage cascades whose retrieval, ranking, and serving components are difficult to replace jointly. We present GRP, a generative recommendation framework that combines retrieval, ranking, and reward modeling in a single encoder-decoder model, and evaluate a progressive path...

Wen-Feng Zhuo, Vincent Xue, Charles Wei et al. · 0 citations
Preprint Sep 2026

OneTrans-V2: Unifying Retrieval, Pre-rank, and Fine-rank with One Transformer in Industrial Recommender

Industrial recommendation systems typically operate as a \emph{cascade} of retrieval, pre-rank, and fine-rank, but these stages are usually trained and served as separate models, causing repeated user-sequence encoding, isolated optimization, and duplicated engineering effort. Building on OneTrans'model-level unificati...

Han-Nan Cao, Jun Guo, Hao-Lei Pei et al. · 0 citations
Book Open access Aug 2026

Tutorial on Generative Recommendation: Foundations and Frontiers

This work comprehensively surveys recent generative recommendation advances through a tri-decoupled perspective, summarize the evolution of tokenization strategies, analyze the trade-offs of major generative architectures, and summarize the transition from supervised next-token prediction to reinforcement-learning-base...

Xiao-Peng Li, Yejing Wang, Hong-Hui Bao et al. · 0 citations
Book Open access Sep 2026

Embedding Subspace Partitioning for Dynamic Multi-Objective Retrieval

Modern industrial recommender systems must optimize across competing objectives, balancing semantic relevance with business metrics such as engagement and revenue. While bi-encoders dominate large-scale retrieval due to their efficiency, they collapse these heterogeneous signals into a single static embedding space. Th...

Shao-Bo Zhang, Alice Leung, Yun-Xiang Ren et al. · 0 citations
Preprint Aug 2026

UniGD: A Unified Generative-Discriminative Framework for Industrial Retrieval

Generative retrieval (GR) is a promising paradigm for industrial search advertising, yet its deployment is constrained by strict relevance and latency requirements. Existing systems cascade GR with an independent relevance model, decoupling the generative likelihood objective from query-ad relevance discrimination, whi...

Shujie Ji, Yawei Kong, Yili Zhao et al. · 0 citations
Preprint Sep 2026

TGR: Advancing Industrial Recommendation from Generative-Paradigm Ranking toward Unified Generation and Reasoning

TGR (Tencent Generative Recommendation), an industrial framework that advances recommendation toward the generative paradigm along three coupled directions, is presented, which is deployed across Tencent production surfaces serving hundreds of millions of users.

Tgr Team Lei Cheng, Hao-Nan Hu, Bei-Bei Kong et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.