Does LLM Focus on the Right Words? Mitigating Context Bias in LLM-based Recommenders

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Bohao, Chen, Jiawei, Liu, Feng, Zhang, Changwang, Wang, Jun, Jin, Canghong, Chen, Chun, Wang, Can
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866908784562012160
author Wang, Bohao
Chen, Jiawei
Liu, Feng
Zhang, Changwang
Wang, Jun
Jin, Canghong
Chen, Chun
Wang, Can
author_facet Wang, Bohao
Chen, Jiawei
Liu, Feng
Zhang, Changwang
Wang, Jun
Jin, Canghong
Chen, Chun
Wang, Can
contents Large language models (LLMs), owing to their extensive open-domain knowledge and semantic reasoning capabilities, have been increasingly integrated into recommender systems (RS). However, a substantial gap remains between the pre-training objectives of LLMs and the specific requirements of recommendation tasks. To address this gap, supervised fine-tuning (SFT) is commonly performed on specially curated recommendation datasets to further enhance their predictive ability. Despite its success, SFT exhibits a critical limitation: it induces Context Bias, whereby the model over-relies on auxiliary tokens, such as task descriptions and prefix-generated tokens, while underutilizing core user interaction tokens that encode user-specific preferences. This bias not only undermines recommendation accuracy but also raises unfairness concerns. To address this issue, we propose Group Distributionally Robust Optimization-based Tuning (GDRT), a novel fine-tuning paradigm that enforces consistent model performance across token groups with varying degrees of relevance to auxiliary tokens. By adaptively upweighting underperforming groups, typically those weakly correlated with auxiliary tokens, GDRT shifts the model's attention from superficial auxiliary cues to informative user interaction tokens, thereby mitigating context bias. Extensive experiments conducted on three public datasets demonstrate that GDRT effectively mitigates context bias, yielding substantial improvements in recommendation accuracy (with an average NDCG@10 gain of 24.29%) and significantly enhancing recommendation fairness. The code is available at https://github.com/WANGBohaO-jpg/GDRT.
format Preprint
id arxiv_https___arxiv_org_abs_2510_10978
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Does LLM Focus on the Right Words? Mitigating Context Bias in LLM-based Recommenders
Wang, Bohao
Chen, Jiawei
Liu, Feng
Zhang, Changwang
Wang, Jun
Jin, Canghong
Chen, Chun
Wang, Can
Information Retrieval
Large language models (LLMs), owing to their extensive open-domain knowledge and semantic reasoning capabilities, have been increasingly integrated into recommender systems (RS). However, a substantial gap remains between the pre-training objectives of LLMs and the specific requirements of recommendation tasks. To address this gap, supervised fine-tuning (SFT) is commonly performed on specially curated recommendation datasets to further enhance their predictive ability. Despite its success, SFT exhibits a critical limitation: it induces Context Bias, whereby the model over-relies on auxiliary tokens, such as task descriptions and prefix-generated tokens, while underutilizing core user interaction tokens that encode user-specific preferences. This bias not only undermines recommendation accuracy but also raises unfairness concerns. To address this issue, we propose Group Distributionally Robust Optimization-based Tuning (GDRT), a novel fine-tuning paradigm that enforces consistent model performance across token groups with varying degrees of relevance to auxiliary tokens. By adaptively upweighting underperforming groups, typically those weakly correlated with auxiliary tokens, GDRT shifts the model's attention from superficial auxiliary cues to informative user interaction tokens, thereby mitigating context bias. Extensive experiments conducted on three public datasets demonstrate that GDRT effectively mitigates context bias, yielding substantial improvements in recommendation accuracy (with an average NDCG@10 gain of 24.29%) and significantly enhancing recommendation fairness. The code is available at https://github.com/WANGBohaO-jpg/GDRT.
title Does LLM Focus on the Right Words? Mitigating Context Bias in LLM-based Recommenders
topic Information Retrieval
url https://arxiv.org/abs/2510.10978