Robust Reinforcement Learning Objectives for Sequential Recommender Systems

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Mozifian, Melissa, Sylvain, Tristan, Evans, Dave, Meng, Lili
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914758667534336
author Mozifian, Melissa
Sylvain, Tristan
Evans, Dave
Meng, Lili
author_facet Mozifian, Melissa
Sylvain, Tristan
Evans, Dave
Meng, Lili
contents Attention-based sequential recommendation methods have shown promise in accurately capturing users' evolving interests from their past interactions. Recent research has also explored the integration of reinforcement learning (RL) into these models, in addition to generating superior user representations. By framing sequential recommendation as an RL problem with reward signals, we can develop recommender systems that incorporate direct user feedback in the form of rewards, enhancing personalization for users. Nonetheless, employing RL algorithms presents challenges, including off-policy training, expansive combinatorial action spaces, and the scarcity of datasets with sufficient reward signals. Contemporary approaches have attempted to combine RL and sequential modeling, incorporating contrastive-based objectives and negative sampling strategies for training the RL component. In this work, we further emphasize the efficacy of contrastive-based objectives paired with augmentation to address datasets with extended horizons. Additionally, we recognize the potential instability issues that may arise during the application of negative sampling. These challenges primarily stem from the data imbalance prevalent in real-world datasets, which is a common issue in offline RL contexts. Furthermore, we introduce an enhanced methodology aimed at providing a more effective solution to these challenges. Experimental results across several real datasets show our method with increased robustness and state-of-the-art performance.
format Preprint
id arxiv_https___arxiv_org_abs_2305_18820
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Robust Reinforcement Learning Objectives for Sequential Recommender Systems
Mozifian, Melissa
Sylvain, Tristan
Evans, Dave
Meng, Lili
Machine Learning
Artificial Intelligence
Information Retrieval
Attention-based sequential recommendation methods have shown promise in accurately capturing users' evolving interests from their past interactions. Recent research has also explored the integration of reinforcement learning (RL) into these models, in addition to generating superior user representations. By framing sequential recommendation as an RL problem with reward signals, we can develop recommender systems that incorporate direct user feedback in the form of rewards, enhancing personalization for users. Nonetheless, employing RL algorithms presents challenges, including off-policy training, expansive combinatorial action spaces, and the scarcity of datasets with sufficient reward signals. Contemporary approaches have attempted to combine RL and sequential modeling, incorporating contrastive-based objectives and negative sampling strategies for training the RL component. In this work, we further emphasize the efficacy of contrastive-based objectives paired with augmentation to address datasets with extended horizons. Additionally, we recognize the potential instability issues that may arise during the application of negative sampling. These challenges primarily stem from the data imbalance prevalent in real-world datasets, which is a common issue in offline RL contexts. Furthermore, we introduce an enhanced methodology aimed at providing a more effective solution to these challenges. Experimental results across several real datasets show our method with increased robustness and state-of-the-art performance.
title Robust Reinforcement Learning Objectives for Sequential Recommender Systems
topic Machine Learning
Artificial Intelligence
Information Retrieval
url https://arxiv.org/abs/2305.18820