Mitigating Preference Leakage via Strict Estimator Separation for Normative Generative Ranking

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Nahhas, Dalia, Cai, Xiaohao, Razzak, Imran, Jameel, Shoaib
Format: Preprint
Publié: 2026
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866917293202604032
author Nahhas, Dalia
Cai, Xiaohao
Razzak, Imran
Jameel, Shoaib
author_facet Nahhas, Dalia
Cai, Xiaohao
Razzak, Imran
Jameel, Shoaib
contents In Generative Information Retrieval (GenIR), the bottleneck has shifted from generation to the selection of candidates, particularly for normative criteria such as cultural relevance. Current LLM-as-a-Judge evaluations often suffer from circularity and preference leakage, where overlapping supervision and evaluation models inflate performance. We address this by formalising cultural relevance as a within-query ranking task and introducing a leakage-free two-judge framework that strictly separates supervision (Judge B) from evaluation (Judge A). On a new benchmark of 33,052 (NGR-33k) culturally grounded stories, we find that while classical baselines yield only modest gains, a dense bi-encoder distilled from a Judge-B-supervised Cross-Encoder is highly effective. Although the Cross-Encoder provides a strong supervision signal for distillation, the distilled BGE-M3 model substantially outperforms it under leakage-free Judge~A evaluation. We validate our framework on the human-curated Moral Stories dataset, showing strong alignment with human norms. Our results demonstrate that rigorous evaluator separation is a prerequisite for credible GenIR evaluation, proving that subtle cultural preferences can be distilled into efficient rankers without leakage.
format Preprint
id arxiv_https___arxiv_org_abs_2602_20800
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Mitigating Preference Leakage via Strict Estimator Separation for Normative Generative Ranking
Nahhas, Dalia
Cai, Xiaohao
Razzak, Imran
Jameel, Shoaib
Information Retrieval
In Generative Information Retrieval (GenIR), the bottleneck has shifted from generation to the selection of candidates, particularly for normative criteria such as cultural relevance. Current LLM-as-a-Judge evaluations often suffer from circularity and preference leakage, where overlapping supervision and evaluation models inflate performance. We address this by formalising cultural relevance as a within-query ranking task and introducing a leakage-free two-judge framework that strictly separates supervision (Judge B) from evaluation (Judge A). On a new benchmark of 33,052 (NGR-33k) culturally grounded stories, we find that while classical baselines yield only modest gains, a dense bi-encoder distilled from a Judge-B-supervised Cross-Encoder is highly effective. Although the Cross-Encoder provides a strong supervision signal for distillation, the distilled BGE-M3 model substantially outperforms it under leakage-free Judge~A evaluation. We validate our framework on the human-curated Moral Stories dataset, showing strong alignment with human norms. Our results demonstrate that rigorous evaluator separation is a prerequisite for credible GenIR evaluation, proving that subtle cultural preferences can be distilled into efficient rankers without leakage.
title Mitigating Preference Leakage via Strict Estimator Separation for Normative Generative Ranking
topic Information Retrieval
url https://arxiv.org/abs/2602.20800