Unichain and Aperiodicity are Sufficient for Asymptotic Optimality of Average-Reward Restless Bandits
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hong, Yige, Xie, Qiaomin, Chen, Yudong, Wang, Weina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Achieving Exponential Asymptotic Optimality in Average-Reward Restless Bandits without Global Attractor Assumption
von: Hong, Yige, et al.
Veröffentlicht: (2024)
von: Hong, Yige, et al.
Veröffentlicht: (2024)
Restless Bandits with Average Reward: Breaking the Uniform Global Attractor Assumption
von: Hong, Yige, et al.
Veröffentlicht: (2023)
von: Hong, Yige, et al.
Veröffentlicht: (2023)
Asymptotically Optimal Policies for Weakly Coupled Markov Decision Processes
von: Goldsztajn, Diego, et al.
Veröffentlicht: (2024)
von: Goldsztajn, Diego, et al.
Veröffentlicht: (2024)
An Optimal-Control Approach to Infinite-Horizon Restless Bandits: Achieving Asymptotic Optimality with Minimal Assumptions
von: YAN, Chen
Veröffentlicht: (2024)
von: YAN, Chen
Veröffentlicht: (2024)
The Gittins index is optimal for dynamic allocation with conditionally independent filtrations
von: Wang, Christopher
Veröffentlicht: (2023)
von: Wang, Christopher
Veröffentlicht: (2023)
Weakly-Coupled Multi-Action Restless Bandits -- Exponential Convergence in Probability
von: Fu, Jing, et al.
Veröffentlicht: (2026)
von: Fu, Jing, et al.
Veröffentlicht: (2026)
Admission Control for A Single Server Waiting Time Process in Heavy Traffic
von: Xie, Bowen, et al.
Veröffentlicht: (2022)
von: Xie, Bowen, et al.
Veröffentlicht: (2022)
Stochastic Control Problems Motivated by Sailboat Trajectory Optimization
von: Ciccarella, Carlo, et al.
Veröffentlicht: (2024)
von: Ciccarella, Carlo, et al.
Veröffentlicht: (2024)
Duality of causal distributionally robust optimization
von: Jiang, Yifan
Veröffentlicht: (2024)
von: Jiang, Yifan
Veröffentlicht: (2024)
Drift Optimization of Regulated Stochastic Models Using Sample Average Approximation
von: Zhou, Zihe, et al.
Veröffentlicht: (2025)
von: Zhou, Zihe, et al.
Veröffentlicht: (2025)
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2024)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2024)
Policy stability and ultimate stationarity in discounted risk-sensitive stochastic control
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
QPLEX Decision Processes: Formulation via Nonlinear Markov Chains and Optimization via Policy Gradients
von: Dieker, Antonius B., et al.
Veröffentlicht: (2026)
von: Dieker, Antonius B., et al.
Veröffentlicht: (2026)
Average Cost Optimality of Partially Observed MDPS: Contraction of Non-linear Filters, Optimal Solutions and Approximations
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2023)
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2023)
Drift Control with Discretionary Stopping for a Diffusion
von: Beneš, Václav E., et al.
Veröffentlicht: (2024)
von: Beneš, Václav E., et al.
Veröffentlicht: (2024)
A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces
von: Kerimkulov, Bekzhan, et al.
Veröffentlicht: (2023)
von: Kerimkulov, Bekzhan, et al.
Veröffentlicht: (2023)
Reduced Sample Complexity in Scenario-Based Control System Design via Constraint Scaling
von: Choi, Jaeseok, et al.
Veröffentlicht: (2024)
von: Choi, Jaeseok, et al.
Veröffentlicht: (2024)
Multi-Action Restless Bandits with Weakly Coupled Constraints: Simultaneous Learning and Control
von: Fu, Jing, et al.
Veröffentlicht: (2024)
von: Fu, Jing, et al.
Veröffentlicht: (2024)
Projection-based Lyapunov method for fully heterogeneous weakly-coupled MDPs
von: Zhang, Xiangcheng, et al.
Veröffentlicht: (2025)
von: Zhang, Xiangcheng, et al.
Veröffentlicht: (2025)
Cost-optimal Management of a Residential Heating System With a Geothermal Energy Storage Under Uncertainty
von: Takam, Paul Honore, et al.
Veröffentlicht: (2025)
von: Takam, Paul Honore, et al.
Veröffentlicht: (2025)
On Strategic Measures and Optimality Properties in Discrete-Time Stochastic Control with Universally Measurable Policies
von: Yu, Huizhen
Veröffentlicht: (2022)
von: Yu, Huizhen
Veröffentlicht: (2022)
Joint Pricing and Innovation Control in Regulated Recycling-Rate Diffusion
von: Xie, Bowen, et al.
Veröffentlicht: (2026)
von: Xie, Bowen, et al.
Veröffentlicht: (2026)
Robust Ergodic Control of Jump-Diffusion Systems under Drift and Intensity Uncertainty
von: Azze, Abel, et al.
Veröffentlicht: (2026)
von: Azze, Abel, et al.
Veröffentlicht: (2026)
Convex Regularization and Convergence of Policy Gradient Flows under Safety Constraints
von: Malo, Pekka, et al.
Veröffentlicht: (2024)
von: Malo, Pekka, et al.
Veröffentlicht: (2024)
Existence of bounded solutions to multiplicative Poisson equations under mixing property
von: Pitera, Marcin, et al.
Veröffentlicht: (2023)
von: Pitera, Marcin, et al.
Veröffentlicht: (2023)
A Note on Stability in Asynchronous Stochastic Approximation without Communication Delays
von: Yu, Huizhen, et al.
Veröffentlicht: (2023)
von: Yu, Huizhen, et al.
Veröffentlicht: (2023)
Reinforcement Learning Methods for the Stochastic Optimal Control of an Industrial Power-to-Heat System
von: Pilling, Eric, et al.
Veröffentlicht: (2024)
von: Pilling, Eric, et al.
Veröffentlicht: (2024)
Single-Item Continuous-Review Inventory Models with Random Supplies
von: Helmes, K. L., et al.
Veröffentlicht: (2024)
von: Helmes, K. L., et al.
Veröffentlicht: (2024)
Continuous-time multi-armed bandits under random intervention times
von: Noba, Kei, et al.
Veröffentlicht: (2026)
von: Noba, Kei, et al.
Veröffentlicht: (2026)
Ergodic Risk Sensitive Control of Diffusions under a General Structural Hypothesis
von: Anugu, Sumith Reddy, et al.
Veröffentlicht: (2025)
von: Anugu, Sumith Reddy, et al.
Veröffentlicht: (2025)
Controlling the low-temperature Ising model using spatiotemporal Markov decision theory
von: de Jongh, M. C., et al.
Veröffentlicht: (2025)
von: de Jongh, M. C., et al.
Veröffentlicht: (2025)
Achieving $\tilde{\mathcal{O}}(1/N)$ Optimality Gap in Restless Bandits through Gaussian Approximation
von: Yan, Chen, et al.
Veröffentlicht: (2024)
von: Yan, Chen, et al.
Veröffentlicht: (2024)
Long run control of nonhomogeneous Markov processes
von: Stettner, Łukasz
Veröffentlicht: (2025)
von: Stettner, Łukasz
Veröffentlicht: (2025)
On Policy Evaluation Algorithms in Distributional Reinforcement Learning
von: Gerstenberg, Julian, et al.
Veröffentlicht: (2024)
von: Gerstenberg, Julian, et al.
Veröffentlicht: (2024)
Optimal strategies for the growth of dual-seeded lattice structures
von: de Jongh, Maike C., et al.
Veröffentlicht: (2025)
von: de Jongh, Maike C., et al.
Veröffentlicht: (2025)
Monotone and Conservative Policy Iteration Beyond the Tabular Case
von: Eshwar, S. R., et al.
Veröffentlicht: (2025)
von: Eshwar, S. R., et al.
Veröffentlicht: (2025)
Large Deviations for Empirical Measures of Self-Interacting Markov Chains
von: Budhiraja, Amarjit, et al.
Veröffentlicht: (2023)
von: Budhiraja, Amarjit, et al.
Veröffentlicht: (2023)
State-Dependent Uncertainty Modeling in Robust Optimal Control Problems through Generalized Semi-Infinite Programming
von: Wehbeh, J., et al.
Veröffentlicht: (2025)
von: Wehbeh, J., et al.
Veröffentlicht: (2025)
Semi-Supervised Clustering of Sparse Graphs: Crossing the Information-Theoretic Threshold
von: Sheng, Junda, et al.
Veröffentlicht: (2022)
von: Sheng, Junda, et al.
Veröffentlicht: (2022)
Utility maximization in multivariate Volterra models
von: Aichinger, Florian, et al.
Veröffentlicht: (2021)
von: Aichinger, Florian, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Achieving Exponential Asymptotic Optimality in Average-Reward Restless Bandits without Global Attractor Assumption
von: Hong, Yige, et al.
Veröffentlicht: (2024) -
Restless Bandits with Average Reward: Breaking the Uniform Global Attractor Assumption
von: Hong, Yige, et al.
Veröffentlicht: (2023) -
Asymptotically Optimal Policies for Weakly Coupled Markov Decision Processes
von: Goldsztajn, Diego, et al.
Veröffentlicht: (2024) -
An Optimal-Control Approach to Infinite-Horizon Restless Bandits: Achieving Asymptotic Optimality with Minimal Assumptions
von: YAN, Chen
Veröffentlicht: (2024) -
The Gittins index is optimal for dynamic allocation with conditionally independent filtrations
von: Wang, Christopher
Veröffentlicht: (2023)