Pessimistic Off-Policy Optimization for Learning to Rank
Fuente:
arXiv
Saved in:
| Main Authors: | Cief, Matej, Kveton, Branislav, Kompan, Michal |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-Validated Off-Policy Evaluation
by: Cief, Matej, et al.
Published: (2024)
by: Cief, Matej, et al.
Published: (2024)
Language-Model Prior Overcomes Cold-Start Items
by: Wang, Shiyu, et al.
Published: (2024)
by: Wang, Shiyu, et al.
Published: (2024)
Context-aware adaptive personalised recommendation: a meta-hybrid
by: Tibensky, Peter, et al.
Published: (2024)
by: Tibensky, Peter, et al.
Published: (2024)
Learning Action Embeddings for Off-Policy Evaluation
by: Cief, Matej, et al.
Published: (2023)
by: Cief, Matej, et al.
Published: (2023)
Ranking Policy Learning via Marketplace Expected Value Estimation From Observational Data
by: Ebrahimzadeh, Ehsan, et al.
Published: (2024)
by: Ebrahimzadeh, Ehsan, et al.
Published: (2024)
Bounded-Abstention Pairwise Learning to Rank
by: Ferrara, Antonio, et al.
Published: (2025)
by: Ferrara, Antonio, et al.
Published: (2025)
A Survey on E-Commerce Learning to Rank
by: Kabir, Md. Ahsanul, et al.
Published: (2024)
by: Kabir, Md. Ahsanul, et al.
Published: (2024)
Decoupled Entity Representation Learning for Pinterest Ads Ranking
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
Bipartite Ranking From Multiple Labels: On Loss Versus Label Aggregation
by: Lukasik, Michal, et al.
Published: (2025)
by: Lukasik, Michal, et al.
Published: (2025)
Breaking the Top-$K$ Barrier: Advancing Top-$K$ Ranking Metrics Optimization in Recommender Systems
by: Yang, Weiqin, et al.
Published: (2025)
by: Yang, Weiqin, et al.
Published: (2025)
Identifiability Matters: Revealing the Hidden Recoverable Condition in Unbiased Learning to Rank
by: Chen, Mouxiang, et al.
Published: (2023)
by: Chen, Mouxiang, et al.
Published: (2023)
Investigating the Robustness of Counterfactual Learning to Rank Models: A Reproducibility Study
by: Niu, Zechun, et al.
Published: (2024)
by: Niu, Zechun, et al.
Published: (2024)
Industry Insights from Comparing Deep Learning and GBDT Models for E-Commerce Learning-to-Rank
by: Lutz, Yunus, et al.
Published: (2025)
by: Lutz, Yunus, et al.
Published: (2025)
Policy-Gradient Training of Language Models for Ranking
by: Gao, Ge, et al.
Published: (2023)
by: Gao, Ge, et al.
Published: (2023)
Localization Boosting for Growth Markets: Mitigating Cross-Locale Behavioral Bias in Learning-to-Rank
by: Seran, Suryaa Veerabathiran, et al.
Published: (2026)
by: Seran, Suryaa Veerabathiran, et al.
Published: (2026)
StealthRank: LLM Ranking Manipulation via Stealthy Prompt Optimization
by: Tang, Yiming, et al.
Published: (2025)
by: Tang, Yiming, et al.
Published: (2025)
Learned Ranking Function: From Short-term Behavior Predictions to Long-term User Satisfaction
by: Wu, Yi, et al.
Published: (2024)
by: Wu, Yi, et al.
Published: (2024)
Distilled Neural Networks for Efficient Learning to Rank
by: Nardini, F. M., et al.
Published: (2022)
by: Nardini, F. M., et al.
Published: (2022)
RRADistill: Distilling LLMs' Passage Ranking Ability for Long-Tail Queries Document Re-Ranking on a Search Engine
by: Choi, Nayoung, et al.
Published: (2024)
by: Choi, Nayoung, et al.
Published: (2024)
Efficient Document Ranking with Learnable Late Interactions
by: Ji, Ziwei, et al.
Published: (2024)
by: Ji, Ziwei, et al.
Published: (2024)
Long Context Modeling with Ranked Memory-Augmented Retrieval
by: Alselwi, Ghadir, et al.
Published: (2025)
by: Alselwi, Ghadir, et al.
Published: (2025)
Pre-trained Recommender Systems: A Causal Debiasing Perspective
by: Lin, Ziqian, et al.
Published: (2023)
by: Lin, Ziqian, et al.
Published: (2023)
Multi-Faceted Large Embedding Tables for Pinterest Ads Ranking
by: Su, Runze, et al.
Published: (2025)
by: Su, Runze, et al.
Published: (2025)
Autoregressive Ranking: Bridging the Gap Between Dual and Cross Encoders
by: Rozonoyer, Benjamin, et al.
Published: (2026)
by: Rozonoyer, Benjamin, et al.
Published: (2026)
RankPO: Preference Optimization for Job-Talent Matching
by: Zhang, Yafei, et al.
Published: (2025)
by: Zhang, Yafei, et al.
Published: (2025)
Async Learned User Embeddings for Ads Delivery Optimization
by: Tang, Mingwei, et al.
Published: (2024)
by: Tang, Mingwei, et al.
Published: (2024)
Rank and Align: Towards Effective Source-free Graph Domain Adaptation
by: Luo, Junyu, et al.
Published: (2024)
by: Luo, Junyu, et al.
Published: (2024)
Ranking Across Different Content Types: The Robust Beauty of Multinomial Blending
by: Lichtenberg, Jan Malte, et al.
Published: (2024)
by: Lichtenberg, Jan Malte, et al.
Published: (2024)
Bridging the Gap: Unpacking the Hidden Challenges in Knowledge Distillation for Online Ranking Systems
by: Khani, Nikhil, et al.
Published: (2024)
by: Khani, Nikhil, et al.
Published: (2024)
GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs
by: Long, Meixiu, et al.
Published: (2025)
by: Long, Meixiu, et al.
Published: (2025)
Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated Judgments
by: Christakopoulou, Evangelia, et al.
Published: (2026)
by: Christakopoulou, Evangelia, et al.
Published: (2026)
RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation
by: Li, Shijun, et al.
Published: (2026)
by: Li, Shijun, et al.
Published: (2026)
A Parameter Update Balancing Algorithm for Multi-task Ranking Models in Recommendation Systems
by: Yuan, Jun, et al.
Published: (2024)
by: Yuan, Jun, et al.
Published: (2024)
Agentic Entropy-Balanced Policy Optimization
by: Dong, Guanting, et al.
Published: (2025)
by: Dong, Guanting, et al.
Published: (2025)
Deep Learning Model Acceleration and Optimization Strategies for Real-Time Recommendation Systems
by: Shao, Junli, et al.
Published: (2025)
by: Shao, Junli, et al.
Published: (2025)
Unifying Ranking and Generation in Query Auto-Completion via Retrieval-Augmented Generation and Multi-Objective Alignment
by: Yuan, Kai, et al.
Published: (2026)
by: Yuan, Kai, et al.
Published: (2026)
Beyond Self-Consistency: Loss-Balanced Perturbation-Based Regularization Improves Industrial-Scale Ads Ranking
by: Ramazanli, Ilqar, et al.
Published: (2025)
by: Ramazanli, Ilqar, et al.
Published: (2025)
Proximal Ranking Policy Optimization for Practical Safety in Counterfactual Learning to Rank
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Enhancing Talent Search Ranking with Role-Aware Expert Mixtures and LLM-based Fine-Grained Job Descriptions
by: Li, Jihang, et al.
Published: (2025)
by: Li, Jihang, et al.
Published: (2025)
AIRwaves at CheckThat! 2025: Retrieving Scientific Sources for Implicit Claims on Social Media with Dual Encoders and Neural Re-Ranking
by: Ashbaugh, Cem, et al.
Published: (2025)
by: Ashbaugh, Cem, et al.
Published: (2025)
Similar Items
-
Cross-Validated Off-Policy Evaluation
by: Cief, Matej, et al.
Published: (2024) -
Language-Model Prior Overcomes Cold-Start Items
by: Wang, Shiyu, et al.
Published: (2024) -
Context-aware adaptive personalised recommendation: a meta-hybrid
by: Tibensky, Peter, et al.
Published: (2024) -
Learning Action Embeddings for Off-Policy Evaluation
by: Cief, Matej, et al.
Published: (2023) -
Ranking Policy Learning via Marketplace Expected Value Estimation From Observational Data
by: Ebrahimzadeh, Ehsan, et al.
Published: (2024)