Bi-Level Optimization for Generative Recommendation: Bridging Tokenization and Generation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bai, Yimeng, Liu, Chang, Zhang, Yang, Wang, Dingxian, Yang, Frank, Rabinovich, Andrew, Rong, Wenge, Feng, Fuli
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918450235965440
author Bai, Yimeng
Liu, Chang
Zhang, Yang
Wang, Dingxian
Yang, Frank
Rabinovich, Andrew
Rong, Wenge
Feng, Fuli
author_facet Bai, Yimeng
Liu, Chang
Zhang, Yang
Wang, Dingxian
Yang, Frank
Rabinovich, Andrew
Rong, Wenge
Feng, Fuli
contents Generative recommendation is emerging as a transformative paradigm by directly generating recommended items, rather than relying on matching. Building such a system typically involves two key components: (1) optimizing the tokenizer to derive suitable item identifiers, and (2) training the recommender based on those identifiers. Existing approaches often treat these components separately--either sequentially or in alternation--overlooking their interdependence. This separation can lead to misalignment: the tokenizer is trained without direct guidance from the recommendation objective, potentially yielding suboptimal identifiers that degrade recommendation performance. To address this, we propose BLOGER, a Bi-Level Optimization for GEnerative Recommendation framework, which explicitly models the interdependence between the tokenizer and the recommender in a unified optimization process. The lower level trains the recommender using tokenized sequences, while the upper level optimizes the tokenizer based on both the tokenization loss and recommendation loss. We adopt a meta-learning approach to solve this bi-level optimization efficiently, and introduce gradient surgery to mitigate gradient conflicts in the upper-level updates, thereby ensuring that item identifiers are both informative and recommendation-aligned. Extensive experiments on multiple real-world datasets demonstrate that BLOGER consistently outperforms state-of-the-art generative recommendation methods while maintaining practical efficiency with no significant additional computational overhead, effectively bridging the gap between item tokenization and autoregressive generation. We release our code at https://github.com/Ten-Mao/BLOGER.
format Preprint
id arxiv_https___arxiv_org_abs_2510_21242
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Bi-Level Optimization for Generative Recommendation: Bridging Tokenization and Generation
Bai, Yimeng
Liu, Chang
Zhang, Yang
Wang, Dingxian
Yang, Frank
Rabinovich, Andrew
Rong, Wenge
Feng, Fuli
Information Retrieval
Generative recommendation is emerging as a transformative paradigm by directly generating recommended items, rather than relying on matching. Building such a system typically involves two key components: (1) optimizing the tokenizer to derive suitable item identifiers, and (2) training the recommender based on those identifiers. Existing approaches often treat these components separately--either sequentially or in alternation--overlooking their interdependence. This separation can lead to misalignment: the tokenizer is trained without direct guidance from the recommendation objective, potentially yielding suboptimal identifiers that degrade recommendation performance. To address this, we propose BLOGER, a Bi-Level Optimization for GEnerative Recommendation framework, which explicitly models the interdependence between the tokenizer and the recommender in a unified optimization process. The lower level trains the recommender using tokenized sequences, while the upper level optimizes the tokenizer based on both the tokenization loss and recommendation loss. We adopt a meta-learning approach to solve this bi-level optimization efficiently, and introduce gradient surgery to mitigate gradient conflicts in the upper-level updates, thereby ensuring that item identifiers are both informative and recommendation-aligned. Extensive experiments on multiple real-world datasets demonstrate that BLOGER consistently outperforms state-of-the-art generative recommendation methods while maintaining practical efficiency with no significant additional computational overhead, effectively bridging the gap between item tokenization and autoregressive generation. We release our code at https://github.com/Ten-Mao/BLOGER.
title Bi-Level Optimization for Generative Recommendation: Bridging Tokenization and Generation
topic Information Retrieval
url https://arxiv.org/abs/2510.21242