Fast Best-in-Class Regret for Contextual Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Girard, Samuel, Bibaut, Aurelien, Gretton, Arthur, Kallus, Nathan, Zenati, Houssam |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Inference after Directionally Stable Adaptive Experiments
by: Shen, Zikai, et al.
Published: (2026)
by: Shen, Zikai, et al.
Published: (2026)
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
by: Shen, Zikai, et al.
Published: (2026)
by: Shen, Zikai, et al.
Published: (2026)
Functional Natural Policy Gradients
by: Bibaut, Aurelien, et al.
Published: (2026)
by: Bibaut, Aurelien, et al.
Published: (2026)
Semiparametric Efficient Bilevel Gradient Estimation
by: Khoury, Fares El, et al.
Published: (2026)
by: Khoury, Fares El, et al.
Published: (2026)
Semiparametric Efficient Test for Interpretable Distributional Treatment Effects
by: Zenati, Houssam, et al.
Published: (2026)
by: Zenati, Houssam, et al.
Published: (2026)
Demistifying Inference after Adaptive Experiments
by: Bibaut, Aurélien, et al.
Published: (2024)
by: Bibaut, Aurélien, et al.
Published: (2024)
Doubly-Robust Estimation of Counterfactual Policy Mean Embeddings
by: Zenati, Houssam, et al.
Published: (2025)
by: Zenati, Houssam, et al.
Published: (2025)
Kernel Treatment Effects with Adaptively Collected Data
by: Zenati, Houssam, et al.
Published: (2025)
by: Zenati, Houssam, et al.
Published: (2025)
Near-Optimal Non-Parametric Sequential Tests and Confidence Sequences with Possibly Dependent Observations
by: Bibaut, Aurelien, et al.
Published: (2022)
by: Bibaut, Aurelien, et al.
Published: (2022)
Simulation-Based Inference for Adaptive Experiments
by: Cho, Brian M, et al.
Published: (2025)
by: Cho, Brian M, et al.
Published: (2025)
Efficient Inference for Inverse Reinforcement Learning and Dynamic Discrete Choice Models
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Nonparametric Instrumental Variable Inference with Many Weak Instruments
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Inverse Reinforcement Learning with Just Classification and a Few Regressions
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Density Ratio-Free Doubly Robust Proxy Causal Learning
by: Bozkurt, Bariscan, et al.
Published: (2025)
by: Bozkurt, Bariscan, et al.
Published: (2025)
Reward Transfer from Inverse Reinforcement Learning: A Coupled Minimax Approach
by: Hao, Guang-Yuan, et al.
Published: (2026)
by: Hao, Guang-Yuan, et al.
Published: (2026)
Automatic Debiased Machine Learning for Smooth Functionals of Nonparametric M-Estimands
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Doubly Robust Proxy Causal Learning with Neural Mean Embeddings
by: Bozkurt, Bariscan, et al.
Published: (2026)
by: Bozkurt, Bariscan, et al.
Published: (2026)
Semiparametric Double Reinforcement Learning with Applications to Long-Term Causal Inference
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
The Value of Personalized Recommendations: Evidence from Netflix
by: Zielnicki, Kevin, et al.
Published: (2025)
by: Zielnicki, Kevin, et al.
Published: (2025)
Multi-Armed Bandits with Interference
by: Jia, Su, et al.
Published: (2024)
by: Jia, Su, et al.
Published: (2024)
Local Anti-Concentration Class: Logarithmic Regret for Greedy Linear Contextual Bandit
by: Kim, Seok-Jin, et al.
Published: (2024)
by: Kim, Seok-Jin, et al.
Published: (2024)
Towards Efficient and Optimal Covariance-Adaptive Algorithms for Combinatorial Semi-Bandits
by: Zhou, Julien, et al.
Published: (2024)
by: Zhou, Julien, et al.
Published: (2024)
Optimal Regret for Policy Optimization in Contextual Bandits
by: Levy, Orin, et al.
Published: (2026)
by: Levy, Orin, et al.
Published: (2026)
Deep Proxy Causal Learning and its Application to Confounded Bandit Policy Evaluation
by: Xu, Liyuan, et al.
Published: (2021)
by: Xu, Liyuan, et al.
Published: (2021)
Inferring the Long-Term Causal Effects of Long-Term Treatments from Short-Term Experiments
by: Tran, Allen, et al.
Published: (2023)
by: Tran, Allen, et al.
Published: (2023)
Nonparametric Jackknife Instrumental Variable Estimation and Confounding Robust Surrogate Indices
by: Bibaut, Aurélien, et al.
Published: (2024)
by: Bibaut, Aurélien, et al.
Published: (2024)
Queue Length Regret Bounds for Contextual Queueing Bandits
by: Bae, Seoungbin, et al.
Published: (2026)
by: Bae, Seoungbin, et al.
Published: (2026)
How Does Variance Shape the Regret in Contextual Bandits?
by: Jia, Zeyu, et al.
Published: (2024)
by: Jia, Zeyu, et al.
Published: (2024)
Improved Best-of-Both-Worlds Regret for Bandits with Delayed Feedback
by: Schlisselberg, Ofir, et al.
Published: (2025)
by: Schlisselberg, Ofir, et al.
Published: (2025)
Active Context Selection Improves Simple Regret in Contextual Bandits
by: Shahverdikondori, Mohammad, et al.
Published: (2026)
by: Shahverdikondori, Mohammad, et al.
Published: (2026)
Variance-Dependent Regret Lower Bounds for Contextual Bandits
by: He, Jiafan, et al.
Published: (2025)
by: He, Jiafan, et al.
Published: (2025)
On the Optimal Regret of Locally Private Linear Contextual Bandit
by: Li, Jiachun, et al.
Published: (2024)
by: Li, Jiachun, et al.
Published: (2024)
Contextual Linear Optimization with Partial Feedback
by: Hu, Yichun, et al.
Published: (2024)
by: Hu, Yichun, et al.
Published: (2024)
Best-of-Both-Worlds Algorithms for Linear Contextual Bandits
by: Kuroki, Yuko, et al.
Published: (2023)
by: Kuroki, Yuko, et al.
Published: (2023)
Double Debiased Machine Learning for Mediation Analysis with Continuous Treatments
by: Zenati, Houssam, et al.
Published: (2025)
by: Zenati, Houssam, et al.
Published: (2025)
Variance-Aware Regret Bounds for Stochastic Contextual Dueling Bandits
by: Di, Qiwei, et al.
Published: (2023)
by: Di, Qiwei, et al.
Published: (2023)
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
by: Kallus, Nathan
Published: (2025)
by: Kallus, Nathan
Published: (2025)
Reward Maximization for Pure Exploration: Minimax Optimal Good Arm Identification for Nonparametric Multi-Armed Bandits
by: Cho, Brian, et al.
Published: (2024)
by: Cho, Brian, et al.
Published: (2024)
Regret Bounds for Adversarial Contextual Bandits with General Function Approximation and Delayed Feedback
by: Levy, Orin, et al.
Published: (2025)
by: Levy, Orin, et al.
Published: (2025)
Fast and Scalable Score-Based Kernel Calibration Tests
by: Glaser, Pierre, et al.
Published: (2025)
by: Glaser, Pierre, et al.
Published: (2025)
Similar Items
-
Efficient Inference after Directionally Stable Adaptive Experiments
by: Shen, Zikai, et al.
Published: (2026) -
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
by: Shen, Zikai, et al.
Published: (2026) -
Functional Natural Policy Gradients
by: Bibaut, Aurelien, et al.
Published: (2026) -
Semiparametric Efficient Bilevel Gradient Estimation
by: Khoury, Fares El, et al.
Published: (2026) -
Semiparametric Efficient Test for Interpretable Distributional Treatment Effects
by: Zenati, Houssam, et al.
Published: (2026)