Model Predictive Control is almost Optimal for Heterogeneous Restless Multi-armed Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Narasimha, Dheeraj, Gast, Nicolas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Model Predictive Control is Almost Optimal for Restless Bandit
by: Gast, Nicolas, et al.
Published: (2024)
by: Gast, Nicolas, et al.
Published: (2024)
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)
by: Avrachenkov, Konstantin, et al.
Published: (2024)
Bandit Allocational Instability
by: Chen, Yilun, et al.
Published: (2026)
by: Chen, Yilun, et al.
Published: (2026)
Achieving Exponential Asymptotic Optimality in Average-Reward Restless Bandits without Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2024)
by: Hong, Yige, et al.
Published: (2024)
Restless Bandits with Average Reward: Breaking the Uniform Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2023)
by: Hong, Yige, et al.
Published: (2023)
Representative Action Selection for Large Action Space Bandit Families
by: Zhou, Quan, et al.
Published: (2025)
by: Zhou, Quan, et al.
Published: (2025)
Multi-Action Restless Bandits with Weakly Coupled Constraints: Simultaneous Learning and Control
by: Fu, Jing, et al.
Published: (2024)
by: Fu, Jing, et al.
Published: (2024)
Representative Action Selection for Large Action Space: From Bandits to MDPs
by: Zhou, Quan, et al.
Published: (2025)
by: Zhou, Quan, et al.
Published: (2025)
Reoptimization Nearly Solves Weakly Coupled Markov Decision Processes
by: Gast, Nicolas, et al.
Published: (2022)
by: Gast, Nicolas, et al.
Published: (2022)
Low-Complexity Algorithm for Restless Bandits with Imperfect Observations
by: Liu, Keqin, et al.
Published: (2021)
by: Liu, Keqin, et al.
Published: (2021)
Stochastic Optimal Control Matching
by: Domingo-Enrich, Carles, et al.
Published: (2023)
by: Domingo-Enrich, Carles, et al.
Published: (2023)
Early Stopping in Contextual Bandits and Inferences
by: Cui, Zihan
Published: (2025)
by: Cui, Zihan
Published: (2025)
Designing Algorithms for Entropic Optimal Transport from an Optimisation Perspective
by: Srinivasan, Vishwak, et al.
Published: (2025)
by: Srinivasan, Vishwak, et al.
Published: (2025)
Computing the Bias of Constant-step Stochastic Approximation with Markovian Noise
by: Allmeier, Sebastian, et al.
Published: (2024)
by: Allmeier, Sebastian, et al.
Published: (2024)
DOPL: Direct Online Preference Learning for Restless Bandits with Preference Feedback
by: Xiong, Guojun, et al.
Published: (2024)
by: Xiong, Guojun, et al.
Published: (2024)
Wasserstein Formulation of Reinforcement Learning. An Optimal Transport Perspective on Policy Optimization
by: Dus, Mathias
Published: (2026)
by: Dus, Mathias
Published: (2026)
Causal Optimal Coupling for Gaussian Input-Output Distributional Data
by: Xu, Daran, et al.
Published: (2026)
by: Xu, Daran, et al.
Published: (2026)
Admission Control of Quasi-Reversible Queueing Systems: Optimization and Reinforcement Learning
by: Comte, Céline, et al.
Published: (2025)
by: Comte, Céline, et al.
Published: (2025)
Controlling the Flow: Stability and Convergence for Stochastic Gradient Descent with Decaying Regularization
by: Kassing, Sebastian, et al.
Published: (2025)
by: Kassing, Sebastian, et al.
Published: (2025)
Convergence of Actor-Critic Learning for Mean Field Games and Mean Field Control in Continuous Spaces
by: Fouque, Jean-Pierre, et al.
Published: (2025)
by: Fouque, Jean-Pierre, et al.
Published: (2025)
Algebraic Reduction of Hidden Markov Models
by: Grigoletto, Tommaso, et al.
Published: (2022)
by: Grigoletto, Tommaso, et al.
Published: (2022)
Markov Kernels, Distances and Optimal Control: A Parable of Linear Quadratic Non-Gaussian Distribution Steering
by: Teter, Alexis M. H., et al.
Published: (2025)
by: Teter, Alexis M. H., et al.
Published: (2025)
Wasserstein Convergence of Score-based Generative Models under Semiconvexity and Discontinuous Gradients
by: Bruno, Stefano, et al.
Published: (2025)
by: Bruno, Stefano, et al.
Published: (2025)
Constrained Density Estimation via Optimal Transport
by: Hu, Yinan, et al.
Published: (2026)
by: Hu, Yinan, et al.
Published: (2026)
Unichain and Aperiodicity are Sufficient for Asymptotic Optimality of Average-Reward Restless Bandits
by: Hong, Yige, et al.
Published: (2024)
by: Hong, Yige, et al.
Published: (2024)
Optimal and instance-dependent guarantees for Markovian linear stochastic approximation
by: Mou, Wenlong, et al.
Published: (2021)
by: Mou, Wenlong, et al.
Published: (2021)
4+3 Phases of Compute-Optimal Neural Scaling Laws
by: Paquette, Elliot, et al.
Published: (2024)
by: Paquette, Elliot, et al.
Published: (2024)
Set Invariance with Probability One for Controlled Diffusion: Score-based Approach
by: Wang, Wenqing, et al.
Published: (2025)
by: Wang, Wenqing, et al.
Published: (2025)
DeepMartingale: Duality of the Optimal Stopping Problem with Expressivity and High-Dimensional Hedging
by: Ye, Junyan, et al.
Published: (2025)
by: Ye, Junyan, et al.
Published: (2025)
Achieving $\tilde{\mathcal{O}}(1/N)$ Optimality Gap in Restless Bandits through Gaussian Approximation
by: Yan, Chen, et al.
Published: (2024)
by: Yan, Chen, et al.
Published: (2024)
Non-convex entropic mean-field optimization via Best Response flow
by: Lascu, Razvan-Andrei, et al.
Published: (2025)
by: Lascu, Razvan-Andrei, et al.
Published: (2025)
ODE approximation for the Adam algorithm: General and overparametrized setting
by: Dereich, Steffen, et al.
Published: (2025)
by: Dereich, Steffen, et al.
Published: (2025)
Flatness-Aware Stochastic Gradient Langevin Dynamics
by: Bruno, Stefano, et al.
Published: (2025)
by: Bruno, Stefano, et al.
Published: (2025)
Benchmarking Diffusion Annealing-Based Bayesian Inverse Problem Solvers
by: Crafts, Evan Scope, et al.
Published: (2025)
by: Crafts, Evan Scope, et al.
Published: (2025)
A stochastic gradient descent algorithm with random search directions
by: Gbaguidi, Eméric
Published: (2025)
by: Gbaguidi, Eméric
Published: (2025)
Algorithmic Stability of Stochastic Gradient Descent with Momentum under Heavy-Tailed Noise
by: Dang, Thanh, et al.
Published: (2025)
by: Dang, Thanh, et al.
Published: (2025)
Reinforcement Learning with Random Time Horizons
by: Borrell, Enric Ribera, et al.
Published: (2025)
by: Borrell, Enric Ribera, et al.
Published: (2025)
A Distributional View of High Dimensional Optimization
by: Benning, Felix
Published: (2025)
by: Benning, Felix
Published: (2025)
Improved Approximation Algorithms for Orthogonally Constrained Problems Using Semidefinite Optimization
by: Cory-Wright, Ryan, et al.
Published: (2025)
by: Cory-Wright, Ryan, et al.
Published: (2025)
When Machine Learning Meets Importance Sampling: A More Efficient Rare Event Estimation Approach
by: Zhao, Ruoning, et al.
Published: (2025)
by: Zhao, Ruoning, et al.
Published: (2025)
Similar Items
-
Model Predictive Control is Almost Optimal for Restless Bandit
by: Gast, Nicolas, et al.
Published: (2024) -
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024) -
Bandit Allocational Instability
by: Chen, Yilun, et al.
Published: (2026) -
Achieving Exponential Asymptotic Optimality in Average-Reward Restless Bandits without Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2024) -
Restless Bandits with Average Reward: Breaking the Uniform Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2023)