Additive Control Variates Dominate Self-Normalisation in Off-Policy Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jeunen, Olivier, Gupta, Shashank |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On (Normalised) Discounted Cumulative Gain as an Off-Policy Evaluation Metric for Top-$n$ Recommendation
von: Jeunen, Olivier, et al.
Veröffentlicht: (2023)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2023)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
von: Gupta, Shashank, et al.
Veröffentlicht: (2024)
von: Gupta, Shashank, et al.
Veröffentlicht: (2024)
Meta Off-Policy Estimation
von: Jeunen, Olivier
Veröffentlicht: (2025)
von: Jeunen, Olivier
Veröffentlicht: (2025)
Unifying On- and Off-Policy Variance Reduction Methods
von: Jeunen, Olivier
Veröffentlicht: (2026)
von: Jeunen, Olivier
Veröffentlicht: (2026)
$Δ\text{-}{\rm OPE}$: Off-Policy Estimation with Pairs of Policies
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
Counterfactual Inference under Thompson Sampling
von: Jeunen, Olivier
Veröffentlicht: (2025)
von: Jeunen, Olivier
Veröffentlicht: (2025)
A Simple Model to Estimate Sharing Effects in Social Networks
von: Jeunen, Olivier
Veröffentlicht: (2024)
von: Jeunen, Olivier
Veröffentlicht: (2024)
Learning Metrics that Maximise Power for Accelerated A/B-Tests
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
Behavioural Effects of Agentic Messaging: A Case Study on a Financial Service Application
von: Jeunen, Olivier, et al.
Veröffentlicht: (2025)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2025)
Multi-Objective Recommendation via Multivariate Policy Learning
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
Proximal Ranking Policy Optimization for Practical Safety in Counterfactual Learning to Rank
von: Gupta, Shashank, et al.
Veröffentlicht: (2024)
von: Gupta, Shashank, et al.
Veröffentlicht: (2024)
Variance Reduction in Ratio Metrics for Efficient Online Experiments
von: Baweja, Shubham, et al.
Veröffentlicht: (2024)
von: Baweja, Shubham, et al.
Veröffentlicht: (2024)
Off-Policy Evaluation and Learning for Matching Markets
von: Hayashi, Yudai, et al.
Veröffentlicht: (2025)
von: Hayashi, Yudai, et al.
Veröffentlicht: (2025)
IntOPE: Off-Policy Evaluation in the Presence of Interference
von: Bai, Yuqi, et al.
Veröffentlicht: (2024)
von: Bai, Yuqi, et al.
Veröffentlicht: (2024)
Agentic Personalisation of Cross-Channel Marketing Experiences
von: Abboud, Sami, et al.
Veröffentlicht: (2025)
von: Abboud, Sami, et al.
Veröffentlicht: (2025)
Practical and Robust Safety Guarantees for Advanced Counterfactual Learning to Rank
von: Gupta, Shashank, et al.
Veröffentlicht: (2024)
von: Gupta, Shashank, et al.
Veröffentlicht: (2024)
Off-policy Evaluation for Payments at Adyen
von: Egg, Alex
Veröffentlicht: (2025)
von: Egg, Alex
Veröffentlicht: (2025)
Pessimistic Off-Policy Optimization for Learning to Rank
von: Cief, Matej, et al.
Veröffentlicht: (2022)
von: Cief, Matej, et al.
Veröffentlicht: (2022)
Learning-to-Rank with Nested Feedback
von: Sagtani, Hitesh, et al.
Veröffentlicht: (2024)
von: Sagtani, Hitesh, et al.
Veröffentlicht: (2024)
Minimizing Live Experiments in Recommender Systems: User Simulation to Evaluate Preference Elicitation Policies
von: Hsu, Chih-Wei, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Wei, et al.
Veröffentlicht: (2024)
Wastewater Pipe Rating Model Using Natural Language Processing
von: Betgeri, Sai Nethra, et al.
Veröffentlicht: (2022)
von: Betgeri, Sai Nethra, et al.
Veröffentlicht: (2022)
Monitoring the Evolution of Behavioural Embeddings in Social Media Recommendation
von: Saket, Srijan, et al.
Veröffentlicht: (2023)
von: Saket, Srijan, et al.
Veröffentlicht: (2023)
Know Your RAG: Dataset Taxonomy and Generation Strategies for Evaluating RAG Systems
von: de Lima, Rafael Teixeira, et al.
Veröffentlicht: (2024)
von: de Lima, Rafael Teixeira, et al.
Veröffentlicht: (2024)
Position Paper: Why the Shooting in the Dark Method Dominates Recommender Systems Practice; A Call to Abandon Anti-Utopian Thinking
von: Rohde, David
Veröffentlicht: (2024)
von: Rohde, David
Veröffentlicht: (2024)
SemSR: Semantics aware robust Session-based Recommendations
von: Narwariya, Jyoti, et al.
Veröffentlicht: (2025)
von: Narwariya, Jyoti, et al.
Veröffentlicht: (2025)
Beyond Retrieval: Ensembling Cross-Encoders and GPT Rerankers with LLMs for Biomedical QA
von: Verma, Shashank, et al.
Veröffentlicht: (2025)
von: Verma, Shashank, et al.
Veröffentlicht: (2025)
Robust Training of Temporal GNNs using Nearest Neighbours based Hard Negatives
von: Gupta, Shubham, et al.
Veröffentlicht: (2024)
von: Gupta, Shubham, et al.
Veröffentlicht: (2024)
Harnessing Feature Clustering For Enhanced Anomaly Detection With Variational Autoencoder And Dynamic Threshold
von: Ale, Tolulope, et al.
Veröffentlicht: (2024)
von: Ale, Tolulope, et al.
Veröffentlicht: (2024)
Entire-Space Variational Information Exploitation for Post-Click Conversion Rate Prediction
von: Fei, Ke, et al.
Veröffentlicht: (2024)
von: Fei, Ke, et al.
Veröffentlicht: (2024)
Leave No One Behind: Online Self-Supervised Self-Distillation for Sequential Recommendation
von: Wei, Shaowei, et al.
Veröffentlicht: (2024)
von: Wei, Shaowei, et al.
Veröffentlicht: (2024)
AskDoc -- Identifying Hidden Healthcare Disparities
von: Gupta, Shashank
Veröffentlicht: (2025)
von: Gupta, Shashank
Veröffentlicht: (2025)
Fast Slate Policy Optimization: Going Beyond Plackett-Luce
von: Sakhi, Otmane, et al.
Veröffentlicht: (2023)
von: Sakhi, Otmane, et al.
Veröffentlicht: (2023)
Guarding Digital Privacy: Exploring User Profiling and Security Enhancements
von: Kohli, Rishika, et al.
Veröffentlicht: (2025)
von: Kohli, Rishika, et al.
Veröffentlicht: (2025)
Simultaneous Unlearning of Multiple Protected User Attributes From Variational Autoencoder Recommenders Using Adversarial Training
von: Escobedo, Gustavo, et al.
Veröffentlicht: (2024)
von: Escobedo, Gustavo, et al.
Veröffentlicht: (2024)
Policy-Guided Causal State Representation for Offline Reinforcement Learning Recommendation
von: Wang, Siyu, et al.
Veröffentlicht: (2025)
von: Wang, Siyu, et al.
Veröffentlicht: (2025)
Understanding the Effects of the Baidu-ULTR Logging Policy on Two-Tower Models
von: de Haan, Morris, et al.
Veröffentlicht: (2024)
von: de Haan, Morris, et al.
Veröffentlicht: (2024)
Powerful A/B-Testing Metrics and Where to Find Them
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
Self-supervised learning for phase retrieval
von: Sechaud, Victor, et al.
Veröffentlicht: (2025)
von: Sechaud, Victor, et al.
Veröffentlicht: (2025)
Efficient and Responsible Adaptation of Large Language Models for Robust and Equitable Top-k Recommendations
von: Kaur, Kirandeep, et al.
Veröffentlicht: (2025)
von: Kaur, Kirandeep, et al.
Veröffentlicht: (2025)
CASP: Support-Aware Offline Policy Selection for Two-Stage Recommender Systems
von: Chapagain, Nilson
Veröffentlicht: (2026)
von: Chapagain, Nilson
Veröffentlicht: (2026)
Ähnliche Einträge
-
On (Normalised) Discounted Cumulative Gain as an Off-Policy Evaluation Metric for Top-$n$ Recommendation
von: Jeunen, Olivier, et al.
Veröffentlicht: (2023) -
Optimal Baseline Corrections for Off-Policy Contextual Bandits
von: Gupta, Shashank, et al.
Veröffentlicht: (2024) -
Meta Off-Policy Estimation
von: Jeunen, Olivier
Veröffentlicht: (2025) -
Unifying On- and Off-Policy Variance Reduction Methods
von: Jeunen, Olivier
Veröffentlicht: (2026) -
$Δ\text{-}{\rm OPE}$: Off-Policy Estimation with Pairs of Policies
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)