Towards Efficient and Optimal Covariance-Adaptive Algorithms for Combinatorial Semi-Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Julien, Gaillard, Pierre, Rahier, Thibaud, Zenati, Houssam, Arbel, Julyan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Logarithmic Regret for Unconstrained Submodular Maximization Stochastic Bandit
von: Zhou, Julien, et al.
Veröffentlicht: (2024)
von: Zhou, Julien, et al.
Veröffentlicht: (2024)
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
von: Boudart, Pierre, et al.
Veröffentlicht: (2025)
von: Boudart, Pierre, et al.
Veröffentlicht: (2025)
Efficient Inference after Directionally Stable Adaptive Experiments
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
Functional Natural Policy Gradients
von: Bibaut, Aurelien, et al.
Veröffentlicht: (2026)
von: Bibaut, Aurelien, et al.
Veröffentlicht: (2026)
Minimax Adaptive Online Nonparametric Regression over Besov Spaces
von: Liautaud, Paul, et al.
Veröffentlicht: (2025)
von: Liautaud, Paul, et al.
Veröffentlicht: (2025)
High-Probability Minimax Adaptive Estimation in Besov Spaces via Online-to-Batch
von: Liautaud, Paul, et al.
Veröffentlicht: (2026)
von: Liautaud, Paul, et al.
Veröffentlicht: (2026)
Minimax Optimal Variance-Aware Regret Bounds for Multinomial Logistic MDPs
von: Boudart, Pierre, et al.
Veröffentlicht: (2026)
von: Boudart, Pierre, et al.
Veröffentlicht: (2026)
Optimal sub-Gaussian variance proxy for 3-mass distributions
von: Atouani, Soufiane, et al.
Veröffentlicht: (2025)
von: Atouani, Soufiane, et al.
Veröffentlicht: (2025)
Federated PCA and Estimation for Spiked Covariance Matrices: Optimal Rates and Efficient Algorithm
von: Li, Jingyang, et al.
Veröffentlicht: (2024)
von: Li, Jingyang, et al.
Veröffentlicht: (2024)
Multimodal Bandits: Regret Lower Bounds and Optimal Algorithms
von: Réveillard, William, et al.
Veröffentlicht: (2025)
von: Réveillard, William, et al.
Veröffentlicht: (2025)
Structured Prediction in Online Learning
von: Boudart, Pierre, et al.
Veröffentlicht: (2024)
von: Boudart, Pierre, et al.
Veröffentlicht: (2024)
On Instability of Minimax Optimal Optimism-Based Bandit Algorithms
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
Statistical Complexity and Optimal Algorithms for Non-linear Ridge Bandits
von: Rajaraman, Nived, et al.
Veröffentlicht: (2023)
von: Rajaraman, Nived, et al.
Veröffentlicht: (2023)
Minimax Rate-Optimal Algorithms for High-Dimensional Stochastic Linear Bandits
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
Nonparametric Instrumental Variable Analysis Without Structural Equations: Debiased Inference on Functionals of Inverse Problems with No Solutions
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
von: Shen, Zikai, et al.
Veröffentlicht: (2026)
Minimax-optimal and Locally-adaptive Online Nonparametric Regression
von: Liautaud, Paul, et al.
Veröffentlicht: (2024)
von: Liautaud, Paul, et al.
Veröffentlicht: (2024)
Optimal Batched Linear Bandits
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
von: Ren, Xuanfei, et al.
Veröffentlicht: (2024)
The Fragility of Optimized Bandit Algorithms
von: Fan, Lin, et al.
Veröffentlicht: (2021)
von: Fan, Lin, et al.
Veröffentlicht: (2021)
Optimal Arm Elimination Algorithms for Combinatorial Bandits
von: Wen, Yuxiao, et al.
Veröffentlicht: (2025)
von: Wen, Yuxiao, et al.
Veröffentlicht: (2025)
Adaptive Smooth Non-Stationary Bandits
von: Suk, Joe
Veröffentlicht: (2024)
von: Suk, Joe
Veröffentlicht: (2024)
On Efficient Estimation of Distributional Treatment Effects under Covariate-Adaptive Randomization
von: Byambadalai, Undral, et al.
Veröffentlicht: (2025)
von: Byambadalai, Undral, et al.
Veröffentlicht: (2025)
Design Experiments to Compare Multi-armed Bandit Algorithms
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
von: Meng, Huiling, et al.
Veröffentlicht: (2026)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
von: Prevost, Adrien, et al.
Veröffentlicht: (2025)
von: Prevost, Adrien, et al.
Veröffentlicht: (2025)
Gaussian Pre-Activations in Neural Networks: Myth or Reality?
von: Wolinski, Pierre, et al.
Veröffentlicht: (2022)
von: Wolinski, Pierre, et al.
Veröffentlicht: (2022)
MetaCURL: Non-stationary Concave Utility Reinforcement Learning
von: Moreno, Bianca Marin, et al.
Veröffentlicht: (2024)
von: Moreno, Bianca Marin, et al.
Veröffentlicht: (2024)
Optimal sub-Gaussian variance proxy for truncated Gaussian and exponential random variables
von: Barreto, Mathias, et al.
Veröffentlicht: (2024)
von: Barreto, Mathias, et al.
Veröffentlicht: (2024)
Minimax and Adaptive Covariance Matrix Estimation under Differential Privacy
von: Cai, T. Tony, et al.
Veröffentlicht: (2026)
von: Cai, T. Tony, et al.
Veröffentlicht: (2026)
Optimal community detection in dense bipartite graphs
von: Chhor, Julien, et al.
Veröffentlicht: (2025)
von: Chhor, Julien, et al.
Veröffentlicht: (2025)
Counterfactual Learning of Stochastic Policies with Continuous Actions
von: Zenati, Houssam, et al.
Veröffentlicht: (2020)
von: Zenati, Houssam, et al.
Veröffentlicht: (2020)
The Adaptivity Barrier in Batched Nonparametric Bandits: Sharp Characterization of the Price of Unknown Margin
von: Jiang, Rong, et al.
Veröffentlicht: (2025)
von: Jiang, Rong, et al.
Veröffentlicht: (2025)
Avoiding the Price of Adaptivity: Inference in Linear Contextual Bandits via Stability
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
von: Praharaj, Samya, et al.
Veröffentlicht: (2025)
Covariates-Adjusted Mixed-Membership Estimation: A Novel Network Model with Optimal Guarantees
von: Fan, Jianqing, et al.
Veröffentlicht: (2025)
von: Fan, Jianqing, et al.
Veröffentlicht: (2025)
A Simple and Optimal Policy Design with Safety against Heavy-Tailed Risk for Stochastic Bandits
von: Simchi-Levi, David, et al.
Veröffentlicht: (2022)
von: Simchi-Levi, David, et al.
Veröffentlicht: (2022)
On the Optimality of Misspecified Spectral Algorithms
von: Zhang, Haobo, et al.
Veröffentlicht: (2023)
von: Zhang, Haobo, et al.
Veröffentlicht: (2023)
Optimal Differentially Private PCA and Estimation for Spiked Covariance Matrices
von: Cai, T. Tony, et al.
Veröffentlicht: (2024)
von: Cai, T. Tony, et al.
Veröffentlicht: (2024)
Choosing the Better Bandit Algorithm under Data Sharing: When Do A/B Experiments Work?
von: Li, Shuangning, et al.
Veröffentlicht: (2025)
von: Li, Shuangning, et al.
Veröffentlicht: (2025)
Stochastic Optimization in Semi-Discrete Optimal Transport: Convergence Analysis and Minimax Rate
von: Genans, Ferdinand, et al.
Veröffentlicht: (2025)
von: Genans, Ferdinand, et al.
Veröffentlicht: (2025)
Self-Distillation is Optimal Among Spectral Shrinkage Estimators in Spiked Covariance Models
von: Lecoiu, Radu, et al.
Veröffentlicht: (2026)
von: Lecoiu, Radu, et al.
Veröffentlicht: (2026)
Minimax-Optimal Spectral Clustering with Covariance Projection for High-Dimensional Anisotropic Mixtures
von: Huang, Chengzhu, et al.
Veröffentlicht: (2025)
von: Huang, Chengzhu, et al.
Veröffentlicht: (2025)
Online Covariance Estimation in Averaged SGD: Improved Batch-Mean Rates and Minimax Optimality via Trajectory Regression
von: Ni, Yijin, et al.
Veröffentlicht: (2026)
von: Ni, Yijin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Logarithmic Regret for Unconstrained Submodular Maximization Stochastic Bandit
von: Zhou, Julien, et al.
Veröffentlicht: (2024) -
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
von: Boudart, Pierre, et al.
Veröffentlicht: (2025) -
Efficient Inference after Directionally Stable Adaptive Experiments
von: Shen, Zikai, et al.
Veröffentlicht: (2026) -
Functional Natural Policy Gradients
von: Bibaut, Aurelien, et al.
Veröffentlicht: (2026) -
Minimax Adaptive Online Nonparametric Regression over Besov Spaces
von: Liautaud, Paul, et al.
Veröffentlicht: (2025)