Lagrangian Relaxation for Multi-Action Partially Observable Restless Bandits: Heuristic Policies and Indexability
Fuente:
arXiv
Saved in:
| Main Authors: | Meshram, Rahul, Kaza, Kesav |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hierarchical Decentralized Stochastic Control for Cyber-Physical Systems
by: Kaza, Kesav, et al.
Published: (2025)
by: Kaza, Kesav, et al.
Published: (2025)
Relaxed Indexability and Index Policy for Partially Observable Restless Bandits
by: Liu, Keqin
Published: (2021)
by: Liu, Keqin
Published: (2021)
Faster Q-Learning Algorithms for Restless Bandits
by: Kakarapalli, Parvish, et al.
Published: (2024)
by: Kakarapalli, Parvish, et al.
Published: (2024)
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
by: Mittal, Vishesh, et al.
Published: (2024)
by: Mittal, Vishesh, et al.
Published: (2024)
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)
by: Avrachenkov, Konstantin, et al.
Published: (2024)
Restless Bandit Problem with Rewards Generated by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2024)
by: Gornet, Jonathan, et al.
Published: (2024)
Risk-Aware Decision Making in Restless Bandits: Theory and Algorithms for Planning and Learning
by: Akbarzadeh, Nima, et al.
Published: (2024)
by: Akbarzadeh, Nima, et al.
Published: (2024)
Action-Conditioned Risk Gating for Safety-Critical Control under Partial Observability
by: Liu, Yushen, et al.
Published: (2026)
by: Liu, Yushen, et al.
Published: (2026)
Mitigating Partial Observability in Adaptive Traffic Signal Control with Transformers
by: Wang, Xiaoyu, et al.
Published: (2024)
by: Wang, Xiaoyu, et al.
Published: (2024)
An LP-based Sampling Policy for Multi-Armed Bandits with Side-Observations and Stochastic Availability
by: Soni, Ashutosh, et al.
Published: (2026)
by: Soni, Ashutosh, et al.
Published: (2026)
Sparse Kalman Identification for Partially Observable Systems via Adaptive Bayesian Learning
by: Mei, Jilan, et al.
Published: (2025)
by: Mei, Jilan, et al.
Published: (2025)
Learning Verifiable Control Policies Using Relaxed Verification
by: Chaudhury, Puja, et al.
Published: (2025)
by: Chaudhury, Puja, et al.
Published: (2025)
Task load dependent decision referrals for joint binary classification in human-automation teams
by: Kaza, Kesav, et al.
Published: (2025)
by: Kaza, Kesav, et al.
Published: (2025)
The Bandit Whisperer: Communication Learning for Restless Bandits
by: Zhao, Yunfan, et al.
Published: (2024)
by: Zhao, Yunfan, et al.
Published: (2024)
Byzantine-Resilient Decentralized Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2023)
by: Zhu, Jingxuan, et al.
Published: (2023)
Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
by: Manupriya, Piyushi, et al.
Published: (2025)
by: Manupriya, Piyushi, et al.
Published: (2025)
Multi-Agent Stage-wise Conservative Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2025)
by: Afsharrad, Amirhossein, et al.
Published: (2025)
Low-Complexity Algorithm for Restless Bandits with Imperfect Observations
by: Liu, Keqin, et al.
Published: (2021)
by: Liu, Keqin, et al.
Published: (2021)
Domain Adaptation of Drag Reduction Policy to Partial Measurements
by: Plaksin, Anton, et al.
Published: (2025)
by: Plaksin, Anton, et al.
Published: (2025)
Learning Decentralized Partially Observable Mean Field Control for Artificial Collective Behavior
by: Cui, Kai, et al.
Published: (2023)
by: Cui, Kai, et al.
Published: (2023)
Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy
by: Juneja, Ishank, et al.
Published: (2026)
by: Juneja, Ishank, et al.
Published: (2026)
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits
by: Zhu, Jingxuan, et al.
Published: (2021)
by: Zhu, Jingxuan, et al.
Published: (2021)
Finite Memory Belief Approximation for Optimal Control in Partially Observable Markov Decision Processes
by: Kim, Mintae
Published: (2026)
by: Kim, Mintae
Published: (2026)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
by: Adams, Katherine B., et al.
Published: (2025)
by: Adams, Katherine B., et al.
Published: (2025)
Model Predictive Control is almost Optimal for Heterogeneous Restless Multi-armed Bandits
by: Narasimha, Dheeraj, et al.
Published: (2025)
by: Narasimha, Dheeraj, et al.
Published: (2025)
Decision Transformer as a Foundation Model for Partially Observable Continuous Control
by: Zhang, Xiangyuan, et al.
Published: (2024)
by: Zhang, Xiangyuan, et al.
Published: (2024)
Model Predictive Control is Almost Optimal for Restless Bandit
by: Gast, Nicolas, et al.
Published: (2024)
by: Gast, Nicolas, et al.
Published: (2024)
Neural Index Policies for Restless Multi-Action Bandits with Heterogeneous Budgets
by: Pandey, Himadri S., et al.
Published: (2025)
by: Pandey, Himadri S., et al.
Published: (2025)
Multi-User mmWave Beam and Rate Adaptation via Combinatorial Satisficing Bandits
by: Özyıldırım, Emre, et al.
Published: (2026)
by: Özyıldırım, Emre, et al.
Published: (2026)
An Intent Modeling and Inference Framework for Autonomous and Remotely Piloted Aerial Systems
by: Kaza, Kesav, et al.
Published: (2024)
by: Kaza, Kesav, et al.
Published: (2024)
DOPL: Direct Online Preference Learning for Restless Bandits with Preference Feedback
by: Xiong, Guojun, et al.
Published: (2024)
by: Xiong, Guojun, et al.
Published: (2024)
Learning to Sparsify Stochastic Linear Bandits
by: Wang, Zhengmiao, et al.
Published: (2026)
by: Wang, Zhengmiao, et al.
Published: (2026)
Finite-Horizon Single-Pull Restless Bandits: An Efficient Index Policy For Scarce Resource Allocation
by: Xiong, Guojun, et al.
Published: (2025)
by: Xiong, Guojun, et al.
Published: (2025)
Bandit Algorithms for Deep Brain Stimulation
by: Gupta, Arkaprava, et al.
Published: (2026)
by: Gupta, Arkaprava, et al.
Published: (2026)
Data-Driven Observability Analysis for Nonlinear Stochastic Systems
by: Massiani, Pierre-François, et al.
Published: (2023)
by: Massiani, Pierre-François, et al.
Published: (2023)
Observability conditions for neural state-space models with eigenvalues and their roots of unity
by: Gracyk, Andrew
Published: (2025)
by: Gracyk, Andrew
Published: (2025)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
by: Choi, Sunmook, et al.
Published: (2025)
by: Choi, Sunmook, et al.
Published: (2025)
Predictive Lagrangian Optimization for Constrained Reinforcement Learning
by: Zhang, Tianqi, et al.
Published: (2025)
by: Zhang, Tianqi, et al.
Published: (2025)
Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks
by: Dai, Yan, et al.
Published: (2024)
by: Dai, Yan, et al.
Published: (2024)
Sparse Mamba: Introducing Controllability, Observability, And Stability To Structural State Space Models
by: Hamdan, Emadeldeen, et al.
Published: (2024)
by: Hamdan, Emadeldeen, et al.
Published: (2024)
Similar Items
-
Hierarchical Decentralized Stochastic Control for Cyber-Physical Systems
by: Kaza, Kesav, et al.
Published: (2025) -
Relaxed Indexability and Index Policy for Partially Observable Restless Bandits
by: Liu, Keqin
Published: (2021) -
Faster Q-Learning Algorithms for Restless Bandits
by: Kakarapalli, Parvish, et al.
Published: (2024) -
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
by: Mittal, Vishesh, et al.
Published: (2024) -
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)