Campaign-2-PT-RAG: LLM-Guided Semantic Product Type Attribution for Scalable Campaign Ranking

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Che, Yiming, Mane, Mansi Ranjit, Gopalakrishnan, Keerthi, Kaghazgaran, Parisa, Dandu, Murali Mohana Krishna, Venkatachalapathy, Archana, Subramaniam, Sinduja, Arora, Yokila, Korpeoglu, Evren, Kumar, Sushant, Achan, Kannan
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918343546503168
author Che, Yiming
Mane, Mansi Ranjit
Gopalakrishnan, Keerthi
Kaghazgaran, Parisa
Dandu, Murali Mohana Krishna
Venkatachalapathy, Archana
Subramaniam, Sinduja
Arora, Yokila
Korpeoglu, Evren
Kumar, Sushant
Achan, Kannan
author_facet Che, Yiming
Mane, Mansi Ranjit
Gopalakrishnan, Keerthi
Kaghazgaran, Parisa
Dandu, Murali Mohana Krishna
Venkatachalapathy, Archana
Subramaniam, Sinduja
Arora, Yokila
Korpeoglu, Evren
Kumar, Sushant
Achan, Kannan
contents E-commerce campaign ranking models require large-scale training labels indicating which users purchased due to campaign influence. However, generating these labels is challenging because campaigns use creative, thematic language that does not directly map to product purchases. Without clear product-level attribution, supervised learning for campaign optimization remains limited. We present Campaign-2-PT-RAG, a scalable label generation framework that constructs user-campaign purchase labels by inferring which product types (PTs) each campaign promotes. The framework first interprets campaign content using large language models (LLMs) to capture implicit intent, then retrieves candidate PTs through semantic search over the platform taxonomy. A structured LLM-based classifier evaluates each PT's relevance, producing a campaign-specific product coverage set. User purchases matching these PTs generate positive training labels for downstream ranking models. This approach reframes the ambiguous attribution problem into a tractable semantic alignment task, enabling scalable and consistent supervision for downstream tasks such as campaign ranking optimization in production e-commerce environments. Experiments on internal and synthetic datasets, validated against expert-annotated campaign-PT mappings, show that our LLM-assisted approach generates high-quality labels with 78-90% precision while maintaining over 99% recall.
format Preprint
id arxiv_https___arxiv_org_abs_2602_10577
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Campaign-2-PT-RAG: LLM-Guided Semantic Product Type Attribution for Scalable Campaign Ranking
Che, Yiming
Mane, Mansi Ranjit
Gopalakrishnan, Keerthi
Kaghazgaran, Parisa
Dandu, Murali Mohana Krishna
Venkatachalapathy, Archana
Subramaniam, Sinduja
Arora, Yokila
Korpeoglu, Evren
Kumar, Sushant
Achan, Kannan
Information Retrieval
E-commerce campaign ranking models require large-scale training labels indicating which users purchased due to campaign influence. However, generating these labels is challenging because campaigns use creative, thematic language that does not directly map to product purchases. Without clear product-level attribution, supervised learning for campaign optimization remains limited. We present Campaign-2-PT-RAG, a scalable label generation framework that constructs user-campaign purchase labels by inferring which product types (PTs) each campaign promotes. The framework first interprets campaign content using large language models (LLMs) to capture implicit intent, then retrieves candidate PTs through semantic search over the platform taxonomy. A structured LLM-based classifier evaluates each PT's relevance, producing a campaign-specific product coverage set. User purchases matching these PTs generate positive training labels for downstream ranking models. This approach reframes the ambiguous attribution problem into a tractable semantic alignment task, enabling scalable and consistent supervision for downstream tasks such as campaign ranking optimization in production e-commerce environments. Experiments on internal and synthetic datasets, validated against expert-annotated campaign-PT mappings, show that our LLM-assisted approach generates high-quality labels with 78-90% precision while maintaining over 99% recall.
title Campaign-2-PT-RAG: LLM-Guided Semantic Product Type Attribution for Scalable Campaign Ranking
topic Information Retrieval
url https://arxiv.org/abs/2602.10577