Tight Sample Complexity Bounds for Entropic Best Policy Identification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Essakine, Amer, Vernade, Claire |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prior-Dependent Allocations for Bayesian Fixed-Budget Best-Arm Identification in Structured Bandits
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2024)
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2024)
Efficient Risk-sensitive Planning via Entropic Risk Measures
von: Marthe, Alexandre, et al.
Veröffentlicht: (2025)
von: Marthe, Alexandre, et al.
Veröffentlicht: (2025)
Commit to the Bit: Reactive Reinforcement Learning Done Right
von: Eberhard, Onno, et al.
Veröffentlicht: (2026)
von: Eberhard, Onno, et al.
Veröffentlicht: (2026)
Partially Observable Reinforcement Learning with Memory Traces
von: Eberhard, Onno, et al.
Veröffentlicht: (2025)
von: Eberhard, Onno, et al.
Veröffentlicht: (2025)
Non-Stationary Lipschitz Bandits
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2025)
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2025)
Variational Bayes Portfolio Construction
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2024)
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2024)
A Pontryagin Perspective on Reinforcement Learning
von: Eberhard, Onno, et al.
Veröffentlicht: (2024)
von: Eberhard, Onno, et al.
Veröffentlicht: (2024)
Split the Differences, Pool the Rest: Provably Efficient Multi-Objective Imitation
von: Sheebaelhamd, Ziyad, et al.
Veröffentlicht: (2026)
von: Sheebaelhamd, Ziyad, et al.
Veröffentlicht: (2026)
Online Decision Deferral under Budget Constraints
von: Reid, Mirabel, et al.
Veröffentlicht: (2024)
von: Reid, Mirabel, et al.
Veröffentlicht: (2024)
Put CASH on Bandits: A Max K-Armed Problem for Automated Machine Learning
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2025)
von: Balef, Amir Rezaei, et al.
Veröffentlicht: (2025)
Quantization-Free Autoregressive Action Transformer
von: Sheebaelhamd, Ziyad, et al.
Veröffentlicht: (2025)
von: Sheebaelhamd, Ziyad, et al.
Veröffentlicht: (2025)
Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model
von: Mortensen, Oliver, et al.
Veröffentlicht: (2025)
von: Mortensen, Oliver, et al.
Veröffentlicht: (2025)
Clustered KL-barycenter design for policy evaluation
von: Weissmann, Simon, et al.
Veröffentlicht: (2025)
von: Weissmann, Simon, et al.
Veröffentlicht: (2025)
Stochastically Constrained Best Arm Identification with Thompson Sampling
von: Yang, Le, et al.
Veröffentlicht: (2025)
von: Yang, Le, et al.
Veröffentlicht: (2025)
On The Complexity of Best-Arm Identification in Non-Stationary Linear Bandits
von: Maynard-Zhang, Leo, et al.
Veröffentlicht: (2026)
von: Maynard-Zhang, Leo, et al.
Veröffentlicht: (2026)
Sample Complexity Bounds for Linear System Identification from a Finite Set
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2024)
von: Chatzikiriakos, Nicolas, et al.
Veröffentlicht: (2024)
Model-Free, Regret-Optimal Best Policy Identification in Online CMDPs
von: Zhou, Zihan, et al.
Veröffentlicht: (2023)
von: Zhou, Zihan, et al.
Veröffentlicht: (2023)
An Optimal Tightness Bound for the Simulation Lemma
von: Lobel, Sam, et al.
Veröffentlicht: (2024)
von: Lobel, Sam, et al.
Veröffentlicht: (2024)
PAC-Bayesian Bounds on Constrained f-Entropic Risk Measures
von: Atbir, Hind, et al.
Veröffentlicht: (2025)
von: Atbir, Hind, et al.
Veröffentlicht: (2025)
The Sample Complexity of Smooth Boosting and the Tightness of the Hardcore Theorem
von: Blanc, Guy, et al.
Veröffentlicht: (2024)
von: Blanc, Guy, et al.
Veröffentlicht: (2024)
Identification and Adaptive Control of Markov Jump Systems: Sample Complexity and Regret Bounds
von: Sattar, Yahya, et al.
Veröffentlicht: (2021)
von: Sattar, Yahya, et al.
Veröffentlicht: (2021)
Which Algorithms Have Tight Generalization Bounds?
von: Gastpar, Michael, et al.
Veröffentlicht: (2024)
von: Gastpar, Michael, et al.
Veröffentlicht: (2024)
Sample Complexity of Causal Identification with Temporal Heterogeneity
von: Rathod, Ameya, et al.
Veröffentlicht: (2026)
von: Rathod, Ameya, et al.
Veröffentlicht: (2026)
Closing the Gap on the Sample Complexity of 1-Identification
von: Li, Zitian, et al.
Veröffentlicht: (2026)
von: Li, Zitian, et al.
Veröffentlicht: (2026)
Fast and Regret Optimal Best Arm Identification: Fundamental Limits and Low-Complexity Algorithms
von: Zhang, Qining, et al.
Veröffentlicht: (2023)
von: Zhang, Qining, et al.
Veröffentlicht: (2023)
Thompson Exploration with Best Challenger Rule in Best Arm Identification
von: Lee, Jongyeong, et al.
Veröffentlicht: (2023)
von: Lee, Jongyeong, et al.
Veröffentlicht: (2023)
Beyond the Lower Bound: Bridging Regret Minimization and Best Arm Identification in Lexicographic Bandits
von: Xue, Bo, et al.
Veröffentlicht: (2025)
von: Xue, Bo, et al.
Veröffentlicht: (2025)
Tight Bounds for Jensen's Gap with Applications to Variational Inference
von: Mazur, Marcin, et al.
Veröffentlicht: (2025)
von: Mazur, Marcin, et al.
Veröffentlicht: (2025)
Achieving $\widetilde{O}(1/ε)$ Sample Complexity for Bilinear Systems Identification under Bounded Noises
von: Yi, Hongyu, et al.
Veröffentlicht: (2026)
von: Yi, Hongyu, et al.
Veröffentlicht: (2026)
Box Thirding: Anytime Best Arm Identification under Insufficient Sampling
von: Hwang, Seohwa, et al.
Veröffentlicht: (2026)
von: Hwang, Seohwa, et al.
Veröffentlicht: (2026)
The Sampling Complexity of Condorcet Winner Identification in Dueling Bandits
von: Saad, El Mehdi, et al.
Veröffentlicht: (2026)
von: Saad, El Mehdi, et al.
Veröffentlicht: (2026)
Reliable Abstention under Adversarial Injections: Tight Lower Bounds and New Upper Bounds
von: Edelman, Ezra, et al.
Veröffentlicht: (2026)
von: Edelman, Ezra, et al.
Veröffentlicht: (2026)
Best Arm Identification with Resource Constraints
von: Li, Zitian, et al.
Veröffentlicht: (2024)
von: Li, Zitian, et al.
Veröffentlicht: (2024)
Optimal Batched Best Arm Identification
von: Jin, Tianyuan, et al.
Veröffentlicht: (2023)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2023)
Cost Aware Best Arm Identification
von: Kanarios, Kellen, et al.
Veröffentlicht: (2024)
von: Kanarios, Kellen, et al.
Veröffentlicht: (2024)
Tight Generalization Bounds for Noiseless Inverse Optimization
von: Fatemi, Pouria, et al.
Veröffentlicht: (2026)
von: Fatemi, Pouria, et al.
Veröffentlicht: (2026)
On The Sample Complexity Bounds In Bilevel Reinforcement Learning
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)
von: Gaur, Mudit, et al.
Veröffentlicht: (2025)
Tight Generalization Bounds for Large-Margin Halfspaces
von: Larsen, Kasper Green, et al.
Veröffentlicht: (2025)
von: Larsen, Kasper Green, et al.
Veröffentlicht: (2025)
Sample Complexity Bounds for Estimating Probability Divergences under Invariances
von: Tahmasebi, Behrooz, et al.
Veröffentlicht: (2023)
von: Tahmasebi, Behrooz, et al.
Veröffentlicht: (2023)
Tight and Efficient Upper Bound on Spectral Norm of Convolutional Layers
von: Grishina, Ekaterina, et al.
Veröffentlicht: (2024)
von: Grishina, Ekaterina, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Prior-Dependent Allocations for Bayesian Fixed-Budget Best-Arm Identification in Structured Bandits
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2024) -
Efficient Risk-sensitive Planning via Entropic Risk Measures
von: Marthe, Alexandre, et al.
Veröffentlicht: (2025) -
Commit to the Bit: Reactive Reinforcement Learning Done Right
von: Eberhard, Onno, et al.
Veröffentlicht: (2026) -
Partially Observable Reinforcement Learning with Memory Traces
von: Eberhard, Onno, et al.
Veröffentlicht: (2025) -
Non-Stationary Lipschitz Bandits
von: Nguyen, Nicolas, et al.
Veröffentlicht: (2025)