Unifying On- and Off-Policy Variance Reduction Methods
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Jeunen, Olivier |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Meta Off-Policy Estimation
von: Jeunen, Olivier
Veröffentlicht: (2025)
von: Jeunen, Olivier
Veröffentlicht: (2025)
Counterfactual Inference under Thompson Sampling
von: Jeunen, Olivier
Veröffentlicht: (2025)
von: Jeunen, Olivier
Veröffentlicht: (2025)
$Δ\text{-}{\rm OPE}$: Off-Policy Estimation with Pairs of Policies
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
Additive Control Variates Dominate Self-Normalisation in Off-Policy Evaluation
von: Jeunen, Olivier, et al.
Veröffentlicht: (2026)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2026)
On (Normalised) Discounted Cumulative Gain as an Off-Policy Evaluation Metric for Top-$n$ Recommendation
von: Jeunen, Olivier, et al.
Veröffentlicht: (2023)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2023)
Variance Reduction in Ratio Metrics for Efficient Online Experiments
von: Baweja, Shubham, et al.
Veröffentlicht: (2024)
von: Baweja, Shubham, et al.
Veröffentlicht: (2024)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
von: Gupta, Shashank, et al.
Veröffentlicht: (2024)
von: Gupta, Shashank, et al.
Veröffentlicht: (2024)
A Simple Model to Estimate Sharing Effects in Social Networks
von: Jeunen, Olivier
Veröffentlicht: (2024)
von: Jeunen, Olivier
Veröffentlicht: (2024)
Learning Metrics that Maximise Power for Accelerated A/B-Tests
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
Behavioural Effects of Agentic Messaging: A Case Study on a Financial Service Application
von: Jeunen, Olivier, et al.
Veröffentlicht: (2025)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2025)
Multi-Objective Recommendation via Multivariate Policy Learning
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024)
Two-stage Risk Control with Application to Ranked Retrieval
von: Xu, Yunpeng, et al.
Veröffentlicht: (2024)
von: Xu, Yunpeng, et al.
Veröffentlicht: (2024)
$t$-Testing the Waters: Empirically Validating Assumptions for Reliable A/B-Testing
von: Jeunen, Olivier
Veröffentlicht: (2025)
von: Jeunen, Olivier
Veröffentlicht: (2025)
Errors in AI-Assisted Retrieval of Medical Literature: A Comparative Study
von: Gao, Jenny, et al.
Veröffentlicht: (2026)
von: Gao, Jenny, et al.
Veröffentlicht: (2026)
dsld: A Socially Relevant Tool for Teaching Statistics
von: Mittal, Aditya, et al.
Veröffentlicht: (2024)
von: Mittal, Aditya, et al.
Veröffentlicht: (2024)
Agentic Personalisation of Cross-Channel Marketing Experiences
von: Abboud, Sami, et al.
Veröffentlicht: (2025)
von: Abboud, Sami, et al.
Veröffentlicht: (2025)
Off-Policy Evaluation and Learning for Matching Markets
von: Hayashi, Yudai, et al.
Veröffentlicht: (2025)
von: Hayashi, Yudai, et al.
Veröffentlicht: (2025)
GACE: Learning Graph-Based Cross-Page Ads Embedding For Click-Through Rate Prediction
von: Wang, Haowen, et al.
Veröffentlicht: (2024)
von: Wang, Haowen, et al.
Veröffentlicht: (2024)
LiDDA: Data Driven Attribution at LinkedIn
von: Bencina, John, et al.
Veröffentlicht: (2025)
von: Bencina, John, et al.
Veröffentlicht: (2025)
Separating and Learning Latent Confounders to Enhancing User Preferences Modeling
von: Xu, Hangtong, et al.
Veröffentlicht: (2023)
von: Xu, Hangtong, et al.
Veröffentlicht: (2023)
Causal Structure Representation Learning of Confounders in Latent Space for Recommendation
von: Xu, Hangtong, et al.
Veröffentlicht: (2023)
von: Xu, Hangtong, et al.
Veröffentlicht: (2023)
Isometry pursuit
von: Koelle, Samson, et al.
Veröffentlicht: (2024)
von: Koelle, Samson, et al.
Veröffentlicht: (2024)
IntOPE: Off-Policy Evaluation in the Presence of Interference
von: Bai, Yuqi, et al.
Veröffentlicht: (2024)
von: Bai, Yuqi, et al.
Veröffentlicht: (2024)
Evaluating the Unseen Capabilities: How Many Theorems Do LLMs Know?
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
Pessimistic Off-Policy Optimization for Learning to Rank
von: Cief, Matej, et al.
Veröffentlicht: (2022)
von: Cief, Matej, et al.
Veröffentlicht: (2022)
Off-policy Evaluation for Payments at Adyen
von: Egg, Alex
Veröffentlicht: (2025)
von: Egg, Alex
Veröffentlicht: (2025)
Learning-to-Rank with Nested Feedback
von: Sagtani, Hitesh, et al.
Veröffentlicht: (2024)
von: Sagtani, Hitesh, et al.
Veröffentlicht: (2024)
Measuring Strategization in Recommendation: Users Adapt Their Behavior to Shape Future Content
von: Cen, Sarah H., et al.
Veröffentlicht: (2024)
von: Cen, Sarah H., et al.
Veröffentlicht: (2024)
Monitoring the Evolution of Behavioural Embeddings in Social Media Recommendation
von: Saket, Srijan, et al.
Veröffentlicht: (2023)
von: Saket, Srijan, et al.
Veröffentlicht: (2023)
Musical Listening Qualia: A Multivariate Approach
von: Mizener, Brendon, et al.
Veröffentlicht: (2024)
von: Mizener, Brendon, et al.
Veröffentlicht: (2024)
A Multifacet Hierarchical Sentiment-Topic Model with Application to Multi-Brand Online Review Analysis
von: Liang, Qiao, et al.
Veröffentlicht: (2025)
von: Liang, Qiao, et al.
Veröffentlicht: (2025)
Scalable recommender system based on factor analysis
von: Ghandwani, Disha, et al.
Veröffentlicht: (2024)
von: Ghandwani, Disha, et al.
Veröffentlicht: (2024)
Seller-Side Experiments under Interference Induced by Feedback Loops in Two-Sided Platforms
von: Zhu, Zhihua, et al.
Veröffentlicht: (2024)
von: Zhu, Zhihua, et al.
Veröffentlicht: (2024)
Logging Policy Design for Off-Policy Evaluation
von: Douglas, Connor, et al.
Veröffentlicht: (2026)
von: Douglas, Connor, et al.
Veröffentlicht: (2026)
Re-ranking Based Diversification: A Unifying View
von: Parambath, Shameem A Puthiya
Veröffentlicht: (2019)
von: Parambath, Shameem A Puthiya
Veröffentlicht: (2019)
nSimplex Zen: A Novel Dimensionality Reduction for Euclidean and Hilbert Spaces
von: Connor, Richard, et al.
Veröffentlicht: (2023)
von: Connor, Richard, et al.
Veröffentlicht: (2023)
UniPinRec: Unifying Generative Retrieval and Ranking at Pinterest Scale
von: Li, Hanyu, et al.
Veröffentlicht: (2026)
von: Li, Hanyu, et al.
Veröffentlicht: (2026)
A Unified Language Model for Large Scale Search, Recommendation, and Reasoning
von: De Nadai, Marco, et al.
Veröffentlicht: (2026)
von: De Nadai, Marco, et al.
Veröffentlicht: (2026)
RankGraph: Unified Heterogeneous Graph Learning for Cross-Domain Recommendation
von: Wu, Renzhi, et al.
Veröffentlicht: (2025)
von: Wu, Renzhi, et al.
Veröffentlicht: (2025)
End-to-End Personalization: Unifying Recommender Systems with Large Language Models
von: Ebrat, Danial, et al.
Veröffentlicht: (2025)
von: Ebrat, Danial, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Meta Off-Policy Estimation
von: Jeunen, Olivier
Veröffentlicht: (2025) -
Counterfactual Inference under Thompson Sampling
von: Jeunen, Olivier
Veröffentlicht: (2025) -
$Δ\text{-}{\rm OPE}$: Off-Policy Estimation with Pairs of Policies
von: Jeunen, Olivier, et al.
Veröffentlicht: (2024) -
Additive Control Variates Dominate Self-Normalisation in Off-Policy Evaluation
von: Jeunen, Olivier, et al.
Veröffentlicht: (2026) -
On (Normalised) Discounted Cumulative Gain as an Off-Policy Evaluation Metric for Top-$n$ Recommendation
von: Jeunen, Olivier, et al.
Veröffentlicht: (2023)