On Universally Optimal Algorithms for A/B Testing
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Po-An, Ariu, Kaito, Proutiere, Alexandre |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Policy Testing in Markov Decision Processes
by: Ariu, Kaito, et al.
Published: (2025)
by: Ariu, Kaito, et al.
Published: (2025)
Revisiting Instance-Optimal Cluster Recovery in the Labeled Stochastic Block Model
by: Ariu, Kaito, et al.
Published: (2023)
by: Ariu, Kaito, et al.
Published: (2023)
Optimal Clustering from Noisy Binary Feedback
by: Ariu, Kaito, et al.
Published: (2019)
by: Ariu, Kaito, et al.
Published: (2019)
Optimal Centered Active Excitation in Linear System Identification
by: Ito, Kaito, et al.
Published: (2026)
by: Ito, Kaito, et al.
Published: (2026)
Best Arm Identification with Fixed Budget: A Large Deviation Perspective
by: Wang, Po-An, et al.
Published: (2023)
by: Wang, Po-An, et al.
Published: (2023)
The Role of Contextual Information in Best Arm Identification
by: Kato, Masahiro, et al.
Published: (2021)
by: Kato, Masahiro, et al.
Published: (2021)
Matroid Semi-Bandits in Sublinear Time
by: Tzeng, Ruo-Chun, et al.
Published: (2024)
by: Tzeng, Ruo-Chun, et al.
Published: (2024)
Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning
by: Stojanovic, Stefan, et al.
Published: (2026)
by: Stojanovic, Stefan, et al.
Published: (2026)
Model-Free Active Exploration in Reinforcement Learning
by: Russo, Alessio, et al.
Published: (2024)
by: Russo, Alessio, et al.
Published: (2024)
Curvature-Guided LoRA: Steering in the pretrained NTK subspace
by: Zheng, Frédéric, et al.
Published: (2026)
by: Zheng, Frédéric, et al.
Published: (2026)
Consensus Group Relative Policy Optimization for Text Generation
by: Ichihara, Yuki, et al.
Published: (2026)
by: Ichihara, Yuki, et al.
Published: (2026)
Conformal Predictions under Markovian Data
by: Zheng, Frédéric, et al.
Published: (2024)
by: Zheng, Frédéric, et al.
Published: (2024)
Learning from Delayed Feedback in Games via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2025)
by: Fujimoto, Yuma, et al.
Published: (2025)
Linear Convergence in Games with Delayed Feedback via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2026)
by: Fujimoto, Yuma, et al.
Published: (2026)
Near-Optimal Clustering in Mixture of Markov Chains
by: Lee, Junghyun, et al.
Published: (2025)
by: Lee, Junghyun, et al.
Published: (2025)
Adaptively Perturbed Mirror Descent for Learning in Games
by: Abe, Kenshi, et al.
Published: (2023)
by: Abe, Kenshi, et al.
Published: (2023)
Shift Before You Learn: Enabling Low-Rank Representations in Reinforcement Learning
by: Dubail, Bastien, et al.
Published: (2025)
by: Dubail, Bastien, et al.
Published: (2025)
Model-free Low-Rank Reinforcement Learning via Leveraged Entry-wise Matrix Estimation
by: Stojanovic, Stefan, et al.
Published: (2024)
by: Stojanovic, Stefan, et al.
Published: (2024)
Return-Aligned Decision Transformer
by: Tanaka, Tsunehiko, et al.
Published: (2024)
by: Tanaka, Tsunehiko, et al.
Published: (2024)
Adversarial Diffusion for Robust Reinforcement Learning
by: Foffano, Daniele, et al.
Published: (2025)
by: Foffano, Daniele, et al.
Published: (2025)
Near-optimal Rank Adaptive Inference of High Dimensional Matrices
by: Zheng, Frédéric, et al.
Published: (2025)
by: Zheng, Frédéric, et al.
Published: (2025)
Adaptive Reinforcement Learning for Unobservable Random Delays
by: Wikman, John, et al.
Published: (2025)
by: Wikman, John, et al.
Published: (2025)
Filtered Direct Preference Optimization
by: Morimura, Tetsuro, et al.
Published: (2024)
by: Morimura, Tetsuro, et al.
Published: (2024)
Low-Rank Bandits via Tight Two-to-Infinity Singular Subspace Recovery
by: Jedra, Yassir, et al.
Published: (2024)
by: Jedra, Yassir, et al.
Published: (2024)
Minimizing Human Intervention in Online Classification
by: Réveillard, William, et al.
Published: (2025)
by: Réveillard, William, et al.
Published: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
by: Foffano, Daniele, et al.
Published: (2023)
by: Foffano, Daniele, et al.
Published: (2023)
Instance-Optimal Estimation with Multiple LLM Judges on a Budget
by: Lee, Junghyun, et al.
Published: (2026)
by: Lee, Junghyun, et al.
Published: (2026)
On the Power of Perturbation under Sampling in Solving Extensive-Form Games
by: Masaka, Wataru, et al.
Published: (2025)
by: Masaka, Wataru, et al.
Published: (2025)
Advantage-Guided Diffusion for Model-Based Reinforcement Learning
by: Foffano, Daniele, et al.
Published: (2026)
by: Foffano, Daniele, et al.
Published: (2026)
Optimal Algorithms for Stochastic Complementary Composite Minimization
by: d'Aspremont, Alexandre, et al.
Published: (2022)
by: d'Aspremont, Alexandre, et al.
Published: (2022)
Caterpillar of Thoughts: The Optimal Test-Time Algorithm for Large Language Models
by: Azarmehr, Amir, et al.
Published: (2026)
by: Azarmehr, Amir, et al.
Published: (2026)
Optimal Algorithms for Augmented Testing of Discrete Distributions
by: Aliakbarpour, Maryam, et al.
Published: (2024)
by: Aliakbarpour, Maryam, et al.
Published: (2024)
The Effect of Optimal Self-Distillation in Noisy Gaussian Mixture Model
by: Takanami, Kaito, et al.
Published: (2025)
by: Takanami, Kaito, et al.
Published: (2025)
Exploiting Similarities in A/B Testing with Off-Policy Estimation
by: Sakhi, Otmane, et al.
Published: (2025)
by: Sakhi, Otmane, et al.
Published: (2025)
Optimal Prediction-Augmented Algorithms for Testing Independence of Distributions
by: Aliakbarpour, Maryam, et al.
Published: (2026)
by: Aliakbarpour, Maryam, et al.
Published: (2026)
Hyperparameter-Free Approach for Faster Minimum Bayes Risk Decoding
by: Jinnai, Yuu, et al.
Published: (2024)
by: Jinnai, Yuu, et al.
Published: (2024)
Bayes correlated equilibria, no-regret dynamics in Bayesian games, and the price of anarchy
by: Fujii, Kaito
Published: (2023)
by: Fujii, Kaito
Published: (2023)
On Uncertainty Quantification for Near-Bayes Optimal Algorithms
by: Wang, Ziyu, et al.
Published: (2024)
by: Wang, Ziyu, et al.
Published: (2024)
Accurate predictive model of band gap with selected important features based on explainable machine learning
by: Lee, Joohwi, et al.
Published: (2025)
by: Lee, Joohwi, et al.
Published: (2025)
Forecast Sports Outcomes under Efficient Market Hypothesis: Theoretical and Experimental Analysis of Odds-Only and Generalised Linear Models
by: Goto, Kaito, et al.
Published: (2026)
by: Goto, Kaito, et al.
Published: (2026)
Similar Items
-
Policy Testing in Markov Decision Processes
by: Ariu, Kaito, et al.
Published: (2025) -
Revisiting Instance-Optimal Cluster Recovery in the Labeled Stochastic Block Model
by: Ariu, Kaito, et al.
Published: (2023) -
Optimal Clustering from Noisy Binary Feedback
by: Ariu, Kaito, et al.
Published: (2019) -
Optimal Centered Active Excitation in Linear System Identification
by: Ito, Kaito, et al.
Published: (2026) -
Best Arm Identification with Fixed Budget: A Large Deviation Perspective
by: Wang, Po-An, et al.
Published: (2023)