DUET: Dual Model Co-Training for Entire Space CTR Prediction

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Xiao, Yutian, Yuan, Meng, Zhuang, Fuzhen, Chen, Wei, Wang, Shukuan, Liu, Shanqi, Feng, Chao, Yu, Wenhui, Li, Xiang, Hu, Lantao, Li, Han, Zhang, Zhao
Format: Preprint
Publié: 2025
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866911236695785472
author Xiao, Yutian
Yuan, Meng
Zhuang, Fuzhen
Chen, Wei
Wang, Shukuan
Liu, Shanqi
Feng, Chao
Yu, Wenhui
Li, Xiang
Hu, Lantao
Li, Han
Zhang, Zhao
author_facet Xiao, Yutian
Yuan, Meng
Zhuang, Fuzhen
Chen, Wei
Wang, Shukuan
Liu, Shanqi
Feng, Chao
Yu, Wenhui
Li, Xiang
Hu, Lantao
Li, Han
Zhang, Zhao
contents The pre-ranking stage plays a pivotal role in large-scale recommender systems but faces an intrinsic trade-off between model expressiveness and computational efficiency. Owing to the massive candidate pool and strict latency constraints, industry systems often rely on lightweight two-tower architectures, which are computationally efficient yet limited in estimation capability. As a result, they struggle to capture the complex synergistic and suppressive relationships among candidate items, which are essential for producing contextually coherent and diverse recommendation lists. Moreover, this simplicity further amplifies the Sample Selection Bias (SSB) problem, as coarse-grained models trained on biased exposure data must generalize to a much larger candidate space with distinct distributions. To address these issues, we propose \textbf{DUET} (\textbf{DU}al Model Co-Training for \textbf{E}ntire Space C\textbf{T}R Prediction), a set-wise pre-ranking framework that achieves expressive modeling under tight computational budgets. Instead of scoring items independently, DUET performs set-level prediction over the entire candidate subset in a single forward pass, enabling information-aware interactions among candidates while amortizing the computational cost across the set. Moreover, a dual model co-training mechanism extends supervision to unexposed items via mutual pseudo-label refinement, effectively mitigating SSB. Validated through extensive offline experiments and online A/B testing, DUET consistently outperforms state-of-the-art baselines and achieves improvements across multiple core business metrics. At present, DUET has been fully deployed in Kuaishou and Kuaishou Lite Apps, serving the main traffic for hundreds of millions of users.
format Preprint
id arxiv_https___arxiv_org_abs_2510_24369
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle DUET: Dual Model Co-Training for Entire Space CTR Prediction
Xiao, Yutian
Yuan, Meng
Zhuang, Fuzhen
Chen, Wei
Wang, Shukuan
Liu, Shanqi
Feng, Chao
Yu, Wenhui
Li, Xiang
Hu, Lantao
Li, Han
Zhang, Zhao
Information Retrieval
The pre-ranking stage plays a pivotal role in large-scale recommender systems but faces an intrinsic trade-off between model expressiveness and computational efficiency. Owing to the massive candidate pool and strict latency constraints, industry systems often rely on lightweight two-tower architectures, which are computationally efficient yet limited in estimation capability. As a result, they struggle to capture the complex synergistic and suppressive relationships among candidate items, which are essential for producing contextually coherent and diverse recommendation lists. Moreover, this simplicity further amplifies the Sample Selection Bias (SSB) problem, as coarse-grained models trained on biased exposure data must generalize to a much larger candidate space with distinct distributions. To address these issues, we propose \textbf{DUET} (\textbf{DU}al Model Co-Training for \textbf{E}ntire Space C\textbf{T}R Prediction), a set-wise pre-ranking framework that achieves expressive modeling under tight computational budgets. Instead of scoring items independently, DUET performs set-level prediction over the entire candidate subset in a single forward pass, enabling information-aware interactions among candidates while amortizing the computational cost across the set. Moreover, a dual model co-training mechanism extends supervision to unexposed items via mutual pseudo-label refinement, effectively mitigating SSB. Validated through extensive offline experiments and online A/B testing, DUET consistently outperforms state-of-the-art baselines and achieves improvements across multiple core business metrics. At present, DUET has been fully deployed in Kuaishou and Kuaishou Lite Apps, serving the main traffic for hundreds of millions of users.
title DUET: Dual Model Co-Training for Entire Space CTR Prediction
topic Information Retrieval
url https://arxiv.org/abs/2510.24369