Powerful A/B-Testing Metrics and Where to Find Them
Fuente:
arXiv
Saved in:
| Main Authors: | Jeunen, Olivier, Baweja, Shubham, Pokharna, Neeti, Ustimenko, Aleksei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Variance Reduction in Ratio Metrics for Efficient Online Experiments
by: Baweja, Shubham, et al.
Published: (2024)
by: Baweja, Shubham, et al.
Published: (2024)
Learning Metrics that Maximise Power for Accelerated A/B-Tests
by: Jeunen, Olivier, et al.
Published: (2024)
by: Jeunen, Olivier, et al.
Published: (2024)
$Δ\text{-}{\rm OPE}$: Off-Policy Estimation with Pairs of Policies
by: Jeunen, Olivier, et al.
Published: (2024)
by: Jeunen, Olivier, et al.
Published: (2024)
On (Normalised) Discounted Cumulative Gain as an Off-Policy Evaluation Metric for Top-$n$ Recommendation
by: Jeunen, Olivier, et al.
Published: (2023)
by: Jeunen, Olivier, et al.
Published: (2023)
Learning-to-Rank with Nested Feedback
by: Sagtani, Hitesh, et al.
Published: (2024)
by: Sagtani, Hitesh, et al.
Published: (2024)
A Simple Model to Estimate Sharing Effects in Social Networks
by: Jeunen, Olivier
Published: (2024)
by: Jeunen, Olivier
Published: (2024)
Multi-Objective Recommendation via Multivariate Policy Learning
by: Jeunen, Olivier, et al.
Published: (2024)
by: Jeunen, Olivier, et al.
Published: (2024)
Meta Off-Policy Estimation
by: Jeunen, Olivier
Published: (2025)
by: Jeunen, Olivier
Published: (2025)
Counterfactual Inference under Thompson Sampling
by: Jeunen, Olivier
Published: (2025)
by: Jeunen, Olivier
Published: (2025)
Unifying On- and Off-Policy Variance Reduction Methods
by: Jeunen, Olivier
Published: (2026)
by: Jeunen, Olivier
Published: (2026)
The Challenge of Using LLMs to Simulate Human Behavior: A Causal Inference Perspective
by: Gui, George, et al.
Published: (2023)
by: Gui, George, et al.
Published: (2023)
Additive Control Variates Dominate Self-Normalisation in Off-Policy Evaluation
by: Jeunen, Olivier, et al.
Published: (2026)
by: Jeunen, Olivier, et al.
Published: (2026)
Recent Advances in Text Analysis
by: Ke, Zheng Tracy, et al.
Published: (2024)
by: Ke, Zheng Tracy, et al.
Published: (2024)
Meta-experiments: Improving experimentation through experimentation
by: Müller, Melanie J. I.
Published: (2024)
by: Müller, Melanie J. I.
Published: (2024)
Comparing how Large Language Models perform against keyword-based searches for social science research data discovery
by: Green, Mark, et al.
Published: (2026)
by: Green, Mark, et al.
Published: (2026)
On Correlating Factors for Domain Adaptation Performance
by: Yuksel, Goksenin, et al.
Published: (2025)
by: Yuksel, Goksenin, et al.
Published: (2025)
LLM-powered Real-time Patent Citation Recommendation for Financial Technologies
by: Deng, Tianang, et al.
Published: (2026)
by: Deng, Tianang, et al.
Published: (2026)
Nationwide Hourly Population Estimating at the Neighborhood Scale in the United States Using Stable-Attendance Anchor Calibration
by: Ning, Huan, et al.
Published: (2026)
by: Ning, Huan, et al.
Published: (2026)
Behavioural Effects of Agentic Messaging: A Case Study on a Financial Service Application
by: Jeunen, Olivier, et al.
Published: (2025)
by: Jeunen, Olivier, et al.
Published: (2025)
Monitoring the Evolution of Behavioural Embeddings in Social Media Recommendation
by: Saket, Srijan, et al.
Published: (2023)
by: Saket, Srijan, et al.
Published: (2023)
Beyond Beats: A Recipe to Song Popularity? A machine learning approach
by: Sebastian, Niklas, et al.
Published: (2024)
by: Sebastian, Niklas, et al.
Published: (2024)
On the Capacity of Distinguishable Synthetic Identity Generation under Face Verification
by: Razeghi, Behrooz
Published: (2026)
by: Razeghi, Behrooz
Published: (2026)
Quantifying Uncertainty in AI Visibility: A Statistical Framework for Generative Search Measurement
by: Sielinski, Ronald
Published: (2026)
by: Sielinski, Ronald
Published: (2026)
hyperFA*IR: A hypergeometric approach to fair rankings with finite candidate pool
by: van Dissel, Mauritz N. Cartier, et al.
Published: (2025)
by: van Dissel, Mauritz N. Cartier, et al.
Published: (2025)
Had enough of experts? Quantitative knowledge retrieval from large language models
by: Selby, David, et al.
Published: (2024)
by: Selby, David, et al.
Published: (2024)
SciConNav: Knowledge navigation through contextual learning of extensive scientific research trajectories
by: Xiang, Shibing, et al.
Published: (2024)
by: Xiang, Shibing, et al.
Published: (2024)
Universal and specific features of Ukrainian economic research: publication analysis based on Crossref data
by: Mryglod, O., et al.
Published: (2021)
by: Mryglod, O., et al.
Published: (2021)
BioDisco: Multi-agent hypothesis generation with dual-mode evidence, iterative feedback and temporal evaluation
by: Ke, Yujing, et al.
Published: (2025)
by: Ke, Yujing, et al.
Published: (2025)
Chronic Stress, Immune Suppression, and Cancer Occurrence: Unveiling the Connection using Survey Data and Predictive Models
by: Lazebnik, Teddy, et al.
Published: (2025)
by: Lazebnik, Teddy, et al.
Published: (2025)
dsld: A Socially Relevant Tool for Teaching Statistics
by: Mittal, Aditya, et al.
Published: (2024)
by: Mittal, Aditya, et al.
Published: (2024)
Errors in AI-Assisted Retrieval of Medical Literature: A Comparative Study
by: Gao, Jenny, et al.
Published: (2026)
by: Gao, Jenny, et al.
Published: (2026)
Analysing the Influence of Macroeconomic Factors on Credit Risk in the UK Banking Sector
by: Sharma, Hemlata, et al.
Published: (2024)
by: Sharma, Hemlata, et al.
Published: (2024)
Network-based Topic Structure Visualization
by: Jeon, Yeseul, et al.
Published: (2024)
by: Jeon, Yeseul, et al.
Published: (2024)
Making Absence Visible: The Roles of Reference and Prompting in Recognizing Missing Information
by: Shoshan, Hagit Ben, et al.
Published: (2026)
by: Shoshan, Hagit Ben, et al.
Published: (2026)
Multi-View Variational Autoencoder for Missing Value Imputation in Untargeted Metabolomics
by: Zhao, Chen, et al.
Published: (2023)
by: Zhao, Chen, et al.
Published: (2023)
Optimal Baseline Corrections for Off-Policy Contextual Bandits
by: Gupta, Shashank, et al.
Published: (2024)
by: Gupta, Shashank, et al.
Published: (2024)
Antibiotic Resistance Microbiology Dataset (ARMD): A Resource for Antimicrobial Resistance from EHRs
by: Haredasht, Fateme Nateghi, et al.
Published: (2025)
by: Haredasht, Fateme Nateghi, et al.
Published: (2025)
Ranking Policy Learning via Marketplace Expected Value Estimation From Observational Data
by: Ebrahimzadeh, Ehsan, et al.
Published: (2024)
by: Ebrahimzadeh, Ehsan, et al.
Published: (2024)
Utilizing citation index and synthetic quality measure to compare Wikipedia languages across various topics
by: Lewoniewski, Włodzimierz, et al.
Published: (2025)
by: Lewoniewski, Włodzimierz, et al.
Published: (2025)
Lightweight Adaptation for LLM-based Technical Service Agent: Latent Logic Augmentation and Robust Noise Reduction
by: Yu, Yi, et al.
Published: (2026)
by: Yu, Yi, et al.
Published: (2026)
Similar Items
-
Variance Reduction in Ratio Metrics for Efficient Online Experiments
by: Baweja, Shubham, et al.
Published: (2024) -
Learning Metrics that Maximise Power for Accelerated A/B-Tests
by: Jeunen, Olivier, et al.
Published: (2024) -
$Δ\text{-}{\rm OPE}$: Off-Policy Estimation with Pairs of Policies
by: Jeunen, Olivier, et al.
Published: (2024) -
On (Normalised) Discounted Cumulative Gain as an Off-Policy Evaluation Metric for Top-$n$ Recommendation
by: Jeunen, Olivier, et al.
Published: (2023) -
Learning-to-Rank with Nested Feedback
by: Sagtani, Hitesh, et al.
Published: (2024)