Risk-Aware Decision Making in Restless Bandits: Theory and Algorithms for Planning and Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Akbarzadeh, Nima, Adulyasak, Yossiri, Delage, Erick |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fair Resource Allocation in Weakly Coupled Markov Decision Processes
by: Tu, Xiaohui, et al.
Published: (2024)
by: Tu, Xiaohui, et al.
Published: (2024)
Navigating Demand Uncertainty in Container Shipping: Deep Reinforcement Learning for Enabling Adaptive and Feasible Master Stowage Planning
by: van Twiller, Jaike, et al.
Published: (2025)
by: van Twiller, Jaike, et al.
Published: (2025)
Faster Q-Learning Algorithms for Restless Bandits
by: Kakarapalli, Parvish, et al.
Published: (2024)
by: Kakarapalli, Parvish, et al.
Published: (2024)
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
by: Mittal, Vishesh, et al.
Published: (2024)
by: Mittal, Vishesh, et al.
Published: (2024)
Planning and Learning in Average Risk-aware MDPs
by: Wang, Weikai, et al.
Published: (2025)
by: Wang, Weikai, et al.
Published: (2025)
Restless Bandit Problem with Rewards Generated by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2024)
by: Gornet, Jonathan, et al.
Published: (2024)
Lagrangian Relaxation for Multi-Action Partially Observable Restless Bandits: Heuristic Policies and Indexability
by: Meshram, Rahul, et al.
Published: (2025)
by: Meshram, Rahul, et al.
Published: (2025)
Low-Complexity Algorithm for Restless Bandits with Imperfect Observations
by: Liu, Keqin, et al.
Published: (2021)
by: Liu, Keqin, et al.
Published: (2021)
A Survey of Contextual Optimization Methods for Decision Making under Uncertainty
by: Sadana, Utsav, et al.
Published: (2023)
by: Sadana, Utsav, et al.
Published: (2023)
The Bandit Whisperer: Communication Learning for Restless Bandits
by: Zhao, Yunfan, et al.
Published: (2024)
by: Zhao, Yunfan, et al.
Published: (2024)
Bandit Algorithms for Deep Brain Stimulation
by: Gupta, Arkaprava, et al.
Published: (2026)
by: Gupta, Arkaprava, et al.
Published: (2026)
Time-Series-Informed Closed-loop Learning for Sequential Decision Making and Control
by: Hirt, Sebastian, et al.
Published: (2024)
by: Hirt, Sebastian, et al.
Published: (2024)
Learning to Sparsify Stochastic Linear Bandits
by: Wang, Zhengmiao, et al.
Published: (2026)
by: Wang, Zhengmiao, et al.
Published: (2026)
DOPL: Direct Online Preference Learning for Restless Bandits with Preference Feedback
by: Xiong, Guojun, et al.
Published: (2024)
by: Xiong, Guojun, et al.
Published: (2024)
Intent-Context Synergy Reinforcement Learning for Autonomous UAV Decision-Making in Air Combat
by: Fu, Jiahao, et al.
Published: (2026)
by: Fu, Jiahao, et al.
Published: (2026)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2021)
by: Zhu, Jingxuan, et al.
Published: (2021)
Mitigating optimistic bias in entropic risk estimation and optimization
by: Sadana, Utsav, et al.
Published: (2024)
by: Sadana, Utsav, et al.
Published: (2024)
Conformal Prediction for Stochastic Decision-Making of PV Power in Electricity Markets
by: Renkema, Yvet, et al.
Published: (2024)
by: Renkema, Yvet, et al.
Published: (2024)
Why Most Optimism Bandit Algorithms Have the Same Regret Analysis: A Simple Unifying Theorem
by: Krishnamurthy, Vikram
Published: (2025)
by: Krishnamurthy, Vikram
Published: (2025)
Model Predictive Control is Almost Optimal for Restless Bandit
by: Gast, Nicolas, et al.
Published: (2024)
by: Gast, Nicolas, et al.
Published: (2024)
Towards a Systems Theory of Algorithms
by: Dörfler, Florian, et al.
Published: (2024)
by: Dörfler, Florian, et al.
Published: (2024)
Deep Learning for Sequential Decision Making under Uncertainty: Foundations, Frameworks, and Frontiers
by: Buyuktahtakin, I. Esra
Published: (2026)
by: Buyuktahtakin, I. Esra
Published: (2026)
Robust Data-driven Prescriptiveness Optimization
by: Poursoltani, Mehran, et al.
Published: (2023)
by: Poursoltani, Mehran, et al.
Published: (2023)
Distributed Online Submodular Maximization under Communication Delays: A Simultaneous Decision-Making Approach
by: Xu, Zirui, et al.
Published: (2026)
by: Xu, Zirui, et al.
Published: (2026)
Risk-averse Decision Making with Contextual Information: Model, Sample Average Approximation, and Kernelization
by: Tao, Yuan, et al.
Published: (2025)
by: Tao, Yuan, et al.
Published: (2025)
Risk-Aware Safe Reinforcement Learning for Control of Stochastic Linear Systems
by: Esmaeili, Babak, et al.
Published: (2025)
by: Esmaeili, Babak, et al.
Published: (2025)
Multi-Objective Reinforcement Learning for Tactical Decision Making for Trucks in Highway Traffic
by: Pathare, Deepthi, et al.
Published: (2026)
by: Pathare, Deepthi, et al.
Published: (2026)
InfraLib: Enabling Reinforcement Learning and Decision-Making for Large-Scale Infrastructure Management
by: Thangeda, Pranay, et al.
Published: (2024)
by: Thangeda, Pranay, et al.
Published: (2024)
Byzantine-Resilient Decentralized Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2023)
by: Zhu, Jingxuan, et al.
Published: (2023)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
by: Choi, Sunmook, et al.
Published: (2025)
by: Choi, Sunmook, et al.
Published: (2025)
EVORA: Deep Evidential Traversability Learning for Risk-Aware Off-Road Autonomy
by: Cai, Xiaoyi, et al.
Published: (2023)
by: Cai, Xiaoyi, et al.
Published: (2023)
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)
by: Avrachenkov, Konstantin, et al.
Published: (2024)
TranSimHub:A Unified Air-Ground Simulation Platform for Multi-Modal Perception and Decision-Making
by: Wang, Maonan, et al.
Published: (2025)
by: Wang, Maonan, et al.
Published: (2025)
Multi-Agent Stage-wise Conservative Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2025)
by: Afsharrad, Amirhossein, et al.
Published: (2025)
Response-Aware Risk-Constrained Control Barrier Function With Application to Vehicles
by: Liao, Qijun, et al.
Published: (2026)
by: Liao, Qijun, et al.
Published: (2026)
A Semi-Supervised Approach for Power System Event Identification
by: Taghipourbazargani, Nima, et al.
Published: (2023)
by: Taghipourbazargani, Nima, et al.
Published: (2023)
Robustifying Conditional Portfolio Decisions via Optimal Transport
by: Nguyen, Viet Anh, et al.
Published: (2021)
by: Nguyen, Viet Anh, et al.
Published: (2021)
A Theory of the Risk for Optimization with Relaxation and its Application to Support Vector Machines
by: Campi, Marco C., et al.
Published: (2020)
by: Campi, Marco C., et al.
Published: (2020)
HPC Application Parameter Autotuning on Edge Devices: A Bandit Learning Approach
by: Hossain, Abrar, et al.
Published: (2025)
by: Hossain, Abrar, et al.
Published: (2025)
Interpolation Conditions for Data Consistency and Prediction in Noisy Linear Systems
by: Vanelli, Martina, et al.
Published: (2025)
by: Vanelli, Martina, et al.
Published: (2025)
Similar Items
-
Fair Resource Allocation in Weakly Coupled Markov Decision Processes
by: Tu, Xiaohui, et al.
Published: (2024) -
Navigating Demand Uncertainty in Container Shipping: Deep Reinforcement Learning for Enabling Adaptive and Feasible Master Stowage Planning
by: van Twiller, Jaike, et al.
Published: (2025) -
Faster Q-Learning Algorithms for Restless Bandits
by: Kakarapalli, Parvish, et al.
Published: (2024) -
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
by: Mittal, Vishesh, et al.
Published: (2024) -
Planning and Learning in Average Risk-aware MDPs
by: Wang, Weikai, et al.
Published: (2025)