Optimal Multi-Fidelity Best-Arm Identification
Fuente:
arXiv
Saved in:
| Main Authors: | Poiani, Riccardo, Degenne, Rémy, Kaufmann, Emilie, Metelli, Alberto Maria, Restelli, Marcello |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Best-Arm Identification in Unimodal Bandits
by: Poiani, Riccardo, et al.
Published: (2024)
by: Poiani, Riccardo, et al.
Published: (2024)
Pure Exploration under Mediators' Feedback
by: Poiani, Riccardo, et al.
Published: (2023)
by: Poiani, Riccardo, et al.
Published: (2023)
Best Arm Identification for Stochastic Rising Bandits
by: Mussi, Marco, et al.
Published: (2023)
by: Mussi, Marco, et al.
Published: (2023)
Truncating Trajectories in Monte Carlo Policy Evaluation: an Adaptive Approach
by: Poiani, Riccardo, et al.
Published: (2024)
by: Poiani, Riccardo, et al.
Published: (2024)
Inverse Reinforcement Learning with Sub-optimal Experts
by: Poiani, Riccardo, et al.
Published: (2024)
by: Poiani, Riccardo, et al.
Published: (2024)
Finding good policies in average-reward Markov Decision Processes without prior knowledge
by: Tuynman, Adrienne, et al.
Published: (2024)
by: Tuynman, Adrienne, et al.
Published: (2024)
Interpetable Target-Feature Aggregation for Multi-Task Learning based on Bias-Variance Analysis
by: Bonetti, Paolo, et al.
Published: (2024)
by: Bonetti, Paolo, et al.
Published: (2024)
Efficient Learning of POMDPs with Known Observation Model in Average-Reward Setting
by: Russo, Alessio, et al.
Published: (2024)
by: Russo, Alessio, et al.
Published: (2024)
A Provably Efficient Option-Based Algorithm for both High-Level and Low-Level Learning
by: Drappo, Gianluca, et al.
Published: (2024)
by: Drappo, Gianluca, et al.
Published: (2024)
Achieving $\widetilde{\mathcal{O}}(\sqrt{T})$ Regret in Average-Reward POMDPs with Known Observation Models
by: Russo, Alessio, et al.
Published: (2025)
by: Russo, Alessio, et al.
Published: (2025)
Projection by Convolution: Optimal Sample Complexity for Reinforcement Learning in Continuous-Space MDPs
by: Maran, Davide, et al.
Published: (2024)
by: Maran, Davide, et al.
Published: (2024)
Local Linearity: the Key for No-regret Reinforcement Learning in Continuous MDPs
by: Maran, Davide, et al.
Published: (2024)
by: Maran, Davide, et al.
Published: (2024)
Policy Gradient with Active Importance Sampling
by: Papini, Matteo, et al.
Published: (2024)
by: Papini, Matteo, et al.
Published: (2024)
Statistical Analysis of Policy Space Compression Problem
by: Molaei, Majid, et al.
Published: (2024)
by: Molaei, Majid, et al.
Published: (2024)
Information Capacity Regret Bounds for Bandits with Mediator Feedback
by: Eldowa, Khaled, et al.
Published: (2024)
by: Eldowa, Khaled, et al.
Published: (2024)
Power Grid Control with Graph-Based Distributed Reinforcement Learning
by: Fabrizio, Carlo, et al.
Published: (2025)
by: Fabrizio, Carlo, et al.
Published: (2025)
Actor-Critic with Active Importance Sampling
by: Molaei, Majid, et al.
Published: (2026)
by: Molaei, Majid, et al.
Published: (2026)
Optimal Batched Best Arm Identification
by: Jin, Tianyuan, et al.
Published: (2023)
by: Jin, Tianyuan, et al.
Published: (2023)
The Batch Complexity of Bandit Pure Exploration
by: Tuynman, Adrienne, et al.
Published: (2025)
by: Tuynman, Adrienne, et al.
Published: (2025)
Scalable Multi-Agent Offline Reinforcement Learning and the Role of Information
by: Zamboni, Riccardo, et al.
Published: (2025)
by: Zamboni, Riccardo, et al.
Published: (2025)
State and Action Factorization in Power Grids
by: Losapio, Gianvito, et al.
Published: (2024)
by: Losapio, Gianvito, et al.
Published: (2024)
Towards Principled Unsupervised Multi-Agent Reinforcement Learning
by: Zamboni, Riccardo, et al.
Published: (2025)
by: Zamboni, Riccardo, et al.
Published: (2025)
Variance-Optimal Arm Selection: Misallocation Minimization and Best Arm Identification
by: Khurshid, Sabrina, et al.
Published: (2025)
by: Khurshid, Sabrina, et al.
Published: (2025)
Bridging Rested and Restless Bandits with Graph-Triggering: Rising and Rotting
by: Genalti, Gianmarco, et al.
Published: (2024)
by: Genalti, Gianmarco, et al.
Published: (2024)
Finite Sample Bounds for Non-Parametric Regression: Optimal Sample Efficiency and Space Complexity
by: Maran, Davide, et al.
Published: (2024)
by: Maran, Davide, et al.
Published: (2024)
Optimal Best Arm Identification under Differential Privacy
by: Jourdan, Marc, et al.
Published: (2025)
by: Jourdan, Marc, et al.
Published: (2025)
Optimal Multi-Objective Best Arm Identification with Fixed Confidence
by: Chen, Zhirui, et al.
Published: (2025)
by: Chen, Zhirui, et al.
Published: (2025)
Nearly Optimal Best Arm Identification for Semiparametric Bandits
by: Kim, Seok-Jin
Published: (2026)
by: Kim, Seok-Jin
Published: (2026)
Minimax and Bayes Optimal Best-Arm Identification
by: Kato, Masahiro
Published: (2025)
by: Kato, Masahiro
Published: (2025)
Near Optimal Best Arm Identification for Clustered Bandits
by: Yash, et al.
Published: (2025)
by: Yash, et al.
Published: (2025)
Autoregressive Bandits
by: Bacchiocchi, Francesco, et al.
Published: (2022)
by: Bacchiocchi, Francesco, et al.
Published: (2022)
Parameterized Projected Bellman Operator
by: Vincent, Théo, et al.
Published: (2023)
by: Vincent, Théo, et al.
Published: (2023)
Optimal Best Arm Identification with Fixed Confidence in Restless Bandits
by: Karthik, P. N., et al.
Published: (2023)
by: Karthik, P. N., et al.
Published: (2023)
Best Arm Identification with Resource Constraints
by: Li, Zitian, et al.
Published: (2024)
by: Li, Zitian, et al.
Published: (2024)
Cost Aware Best Arm Identification
by: Kanarios, Kellen, et al.
Published: (2024)
by: Kanarios, Kellen, et al.
Published: (2024)
Suboptimal Performance of the Bayes Optimal Algorithm in Frequentist Best Arm Identification
by: Komiyama, Junpei
Published: (2022)
by: Komiyama, Junpei
Published: (2022)
Bandit Pareto Set Identification in a Multi-Output Linear Model
by: Kone, Cyrille, et al.
Published: (2025)
by: Kone, Cyrille, et al.
Published: (2025)
Thompson Exploration with Best Challenger Rule in Best Arm Identification
by: Lee, Jongyeong, et al.
Published: (2023)
by: Lee, Jongyeong, et al.
Published: (2023)
Optimal Top-Two Method for Best Arm Identification and Fluid Analysis
by: Bandyopadhyay, Agniv, et al.
Published: (2024)
by: Bandyopadhyay, Agniv, et al.
Published: (2024)
Optimal Best-Arm Identification under Fixed Confidence with Multiple Optima
by: Truong, Lan V.
Published: (2025)
by: Truong, Lan V.
Published: (2025)
Similar Items
-
Best-Arm Identification in Unimodal Bandits
by: Poiani, Riccardo, et al.
Published: (2024) -
Pure Exploration under Mediators' Feedback
by: Poiani, Riccardo, et al.
Published: (2023) -
Best Arm Identification for Stochastic Rising Bandits
by: Mussi, Marco, et al.
Published: (2023) -
Truncating Trajectories in Monte Carlo Policy Evaluation: an Adaptive Approach
by: Poiani, Riccardo, et al.
Published: (2024) -
Inverse Reinforcement Learning with Sub-optimal Experts
by: Poiani, Riccardo, et al.
Published: (2024)