Generative Retrieval with Semantic Tree-Structured Item Identifiers via Contrastive Learning

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Si, Zihua, Sun, Zhongxiang, Chen, Jiale, Chen, Guozhang, Zang, Xiaoxue, Zheng, Kai, Song, Yang, Zhang, Xiao, Xu, Jun, Gai, Kun
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917714301288448
author Si, Zihua
Sun, Zhongxiang
Chen, Jiale
Chen, Guozhang
Zang, Xiaoxue
Zheng, Kai
Song, Yang
Zhang, Xiao
Xu, Jun
Gai, Kun
author_facet Si, Zihua
Sun, Zhongxiang
Chen, Jiale
Chen, Guozhang
Zang, Xiaoxue
Zheng, Kai
Song, Yang
Zhang, Xiao
Xu, Jun
Gai, Kun
contents The retrieval phase is a vital component in recommendation systems, requiring the model to be effective and efficient. Recently, generative retrieval has become an emerging paradigm for document retrieval, showing notable performance. These methods enjoy merits like being end-to-end differentiable, suggesting their viability in recommendation. However, these methods fall short in efficiency and effectiveness for large-scale recommendations. To obtain efficiency and effectiveness, this paper introduces a generative retrieval framework, namely SEATER, which learns SEmAntic Tree-structured item identifiERs via contrastive learning. Specifically, we employ an encoder-decoder model to extract user interests from historical behaviors and retrieve candidates via tree-structured item identifiers. SEATER devises a balanced k-ary tree structure of item identifiers, allocating semantic space to each token individually. This strategy maintains semantic consistency within the same level, while distinct levels correlate to varying semantic granularities. This structure also maintains consistent and fast inference speed for all items. Considering the tree structure, SEATER learns identifier tokens' semantics, hierarchical relationships, and inter-token dependencies. To achieve this, we incorporate two contrastive learning tasks with the generation task to optimize both the model and identifiers. The infoNCE loss aligns the token embeddings based on their hierarchical positions. The triplet loss ranks similar identifiers in desired orders. In this way, SEATER achieves both efficiency and effectiveness. Extensive experiments on three public datasets and an industrial dataset have demonstrated that SEATER outperforms state-of-the-art models significantly.
format Preprint
id arxiv_https___arxiv_org_abs_2309_13375
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Generative Retrieval with Semantic Tree-Structured Item Identifiers via Contrastive Learning
Si, Zihua
Sun, Zhongxiang
Chen, Jiale
Chen, Guozhang
Zang, Xiaoxue
Zheng, Kai
Song, Yang
Zhang, Xiao
Xu, Jun
Gai, Kun
Information Retrieval
The retrieval phase is a vital component in recommendation systems, requiring the model to be effective and efficient. Recently, generative retrieval has become an emerging paradigm for document retrieval, showing notable performance. These methods enjoy merits like being end-to-end differentiable, suggesting their viability in recommendation. However, these methods fall short in efficiency and effectiveness for large-scale recommendations. To obtain efficiency and effectiveness, this paper introduces a generative retrieval framework, namely SEATER, which learns SEmAntic Tree-structured item identifiERs via contrastive learning. Specifically, we employ an encoder-decoder model to extract user interests from historical behaviors and retrieve candidates via tree-structured item identifiers. SEATER devises a balanced k-ary tree structure of item identifiers, allocating semantic space to each token individually. This strategy maintains semantic consistency within the same level, while distinct levels correlate to varying semantic granularities. This structure also maintains consistent and fast inference speed for all items. Considering the tree structure, SEATER learns identifier tokens' semantics, hierarchical relationships, and inter-token dependencies. To achieve this, we incorporate two contrastive learning tasks with the generation task to optimize both the model and identifiers. The infoNCE loss aligns the token embeddings based on their hierarchical positions. The triplet loss ranks similar identifiers in desired orders. In this way, SEATER achieves both efficiency and effectiveness. Extensive experiments on three public datasets and an industrial dataset have demonstrated that SEATER outperforms state-of-the-art models significantly.
title Generative Retrieval with Semantic Tree-Structured Item Identifiers via Contrastive Learning
topic Information Retrieval
url https://arxiv.org/abs/2309.13375