Restless Bandit Problem with Rewards Generated by a Linear Gaussian Dynamical System
Fuente:
arXiv
Saved in:
| Main Authors: | Gornet, Jonathan, Sinopoli, Bruno |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Exploration-free Method for a Linear Stochastic Bandit Driven by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2025)
by: Gornet, Jonathan, et al.
Published: (2025)
An Adaptive Method for Contextual Stochastic Multi-armed Bandits with Rewards Generated by a Linear Dynamical System
by: Gornet, Jonathan, et al.
Published: (2024)
by: Gornet, Jonathan, et al.
Published: (2024)
A Control Theory inspired Exploration Method for a Linear Bandit driven by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2025)
by: Gornet, Jonathan, et al.
Published: (2025)
HyperController: A Hyperparameter Controller for Fast and Stable Training of Reinforcement Learning Neural Networks
by: Gornet, Jonathan, et al.
Published: (2025)
by: Gornet, Jonathan, et al.
Published: (2025)
A Data-Integrated Framework for Learning Fractional-Order Nonlinear Dynamical Systems
by: Yaghooti, Bahram, et al.
Published: (2025)
by: Yaghooti, Bahram, et al.
Published: (2025)
Faster Q-Learning Algorithms for Restless Bandits
by: Kakarapalli, Parvish, et al.
Published: (2024)
by: Kakarapalli, Parvish, et al.
Published: (2024)
Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes
by: Mittal, Vishesh, et al.
Published: (2024)
by: Mittal, Vishesh, et al.
Published: (2024)
Risk-Aware Decision Making in Restless Bandits: Theory and Algorithms for Planning and Learning
by: Akbarzadeh, Nima, et al.
Published: (2024)
by: Akbarzadeh, Nima, et al.
Published: (2024)
Lagrangian Relaxation for Multi-Action Partially Observable Restless Bandits: Heuristic Policies and Indexability
by: Meshram, Rahul, et al.
Published: (2025)
by: Meshram, Rahul, et al.
Published: (2025)
Model-Free Learning and Optimal Policy Design in Multi-Agent MDPs Under Probabilistic Agent Dropout
by: Fiscko, Carmel, et al.
Published: (2023)
by: Fiscko, Carmel, et al.
Published: (2023)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
by: Choi, Sunmook, et al.
Published: (2025)
by: Choi, Sunmook, et al.
Published: (2025)
System Identification for Continuous-time Linear Dynamical Systems
by: Halmos, Peter, et al.
Published: (2023)
by: Halmos, Peter, et al.
Published: (2023)
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)
by: Avrachenkov, Konstantin, et al.
Published: (2024)
Learning of Linear Dynamical Systems as a Non-Commutative Polynomial Optimization Problem
by: Zhou, Quan, et al.
Published: (2020)
by: Zhou, Quan, et al.
Published: (2020)
Learning to Sparsify Stochastic Linear Bandits
by: Wang, Zhengmiao, et al.
Published: (2026)
by: Wang, Zhengmiao, et al.
Published: (2026)
Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
by: Manupriya, Piyushi, et al.
Published: (2025)
by: Manupriya, Piyushi, et al.
Published: (2025)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
by: Adams, Katherine B., et al.
Published: (2025)
by: Adams, Katherine B., et al.
Published: (2025)
Multi-Agent Stage-wise Conservative Linear Bandits
by: Afsharrad, Amirhossein, et al.
Published: (2025)
by: Afsharrad, Amirhossein, et al.
Published: (2025)
The Bandit Whisperer: Communication Learning for Restless Bandits
by: Zhao, Yunfan, et al.
Published: (2024)
by: Zhao, Yunfan, et al.
Published: (2024)
Balancing Act: Prioritization Strategies for LLM-Designed Restless Bandit Rewards
by: Verma, Shresth, et al.
Published: (2024)
by: Verma, Shresth, et al.
Published: (2024)
Tangential Randomization in Linear Bandits (TRAiL): Guaranteed Inference and Regret Bounds
by: Güçlü, Arda, et al.
Published: (2024)
by: Güçlü, Arda, et al.
Published: (2024)
A New Approach to Controlling Linear Dynamical Systems
by: Brahmbhatt, Anand, et al.
Published: (2025)
by: Brahmbhatt, Anand, et al.
Published: (2025)
Interpretable Physics Extraction from Data for Linear Dynamical Systems using Lie Generator Networks
by: Jamil, Shafayeth, et al.
Published: (2026)
by: Jamil, Shafayeth, et al.
Published: (2026)
Switched Linear Ensemble Systems and Structural Controllability
by: Yin, Haoyu, et al.
Published: (2025)
by: Yin, Haoyu, et al.
Published: (2025)
Low-Complexity Algorithm for Restless Bandits with Imperfect Observations
by: Liu, Keqin, et al.
Published: (2021)
by: Liu, Keqin, et al.
Published: (2021)
Joint Learning of Linear Time-Invariant Dynamical Systems
by: Modi, Aditya, et al.
Published: (2021)
by: Modi, Aditya, et al.
Published: (2021)
Symmetric Linear Dynamical Systems are Learnable from Few Observations
by: Vu, Minh, et al.
Published: (2025)
by: Vu, Minh, et al.
Published: (2025)
Efficient Spectral Control of Partially Observed Linear Dynamical Systems
by: Brahmbhatt, Anand, et al.
Published: (2025)
by: Brahmbhatt, Anand, et al.
Published: (2025)
A PAC-Bayes Approach for Controlling Unknown Linear Discrete-time Systems
by: Luo, Yujia, et al.
Published: (2026)
by: Luo, Yujia, et al.
Published: (2026)
Model Predictive Control is Almost Optimal for Restless Bandit
by: Gast, Nicolas, et al.
Published: (2024)
by: Gast, Nicolas, et al.
Published: (2024)
Physics-informed Gaussian Processes as Linear Model Predictive Controller
by: Tebbe, Jörn, et al.
Published: (2024)
by: Tebbe, Jörn, et al.
Published: (2024)
On Zero-sum Game Representation for Replicator Dynamics
by: Yin, Haoyu, et al.
Published: (2025)
by: Yin, Haoyu, et al.
Published: (2025)
On Permanence of Conservative Replicator Dynamics with Four Strategies
by: Yin, Haoyu, et al.
Published: (2026)
by: Yin, Haoyu, et al.
Published: (2026)
Asynchronous Distributed Gaussian Process Regression for Online Learning and Dynamical Systems: Complementary Document
by: Yang, Zewen, et al.
Published: (2024)
by: Yang, Zewen, et al.
Published: (2024)
High Effort, Low Gain: Fundamental Limits of Active Learning for Linear Dynamical Systems
by: Chatzikiriakos, Nicolas, et al.
Published: (2025)
by: Chatzikiriakos, Nicolas, et al.
Published: (2025)
Identifying Large-Scale Linear Parameter Varying Systems with Dynamic Mode Decomposition Methods
by: Jordanou, Jean Panaioti, et al.
Published: (2025)
by: Jordanou, Jean Panaioti, et al.
Published: (2025)
DOPL: Direct Online Preference Learning for Restless Bandits with Preference Feedback
by: Xiong, Guojun, et al.
Published: (2024)
by: Xiong, Guojun, et al.
Published: (2024)
Restless Bandits with Average Reward: Breaking the Uniform Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2023)
by: Hong, Yige, et al.
Published: (2023)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part I
by: Tian, Yi, et al.
Published: (2022)
by: Tian, Yi, et al.
Published: (2022)
Cost-Driven Representation Learning for Linear Quadratic Gaussian Control: Part II
by: Tian, Yi, et al.
Published: (2026)
by: Tian, Yi, et al.
Published: (2026)
Similar Items
-
An Exploration-free Method for a Linear Stochastic Bandit Driven by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2025) -
An Adaptive Method for Contextual Stochastic Multi-armed Bandits with Rewards Generated by a Linear Dynamical System
by: Gornet, Jonathan, et al.
Published: (2024) -
A Control Theory inspired Exploration Method for a Linear Bandit driven by a Linear Gaussian Dynamical System
by: Gornet, Jonathan, et al.
Published: (2025) -
HyperController: A Hyperparameter Controller for Fast and Stable Training of Reinforcement Learning Neural Networks
by: Gornet, Jonathan, et al.
Published: (2025) -
A Data-Integrated Framework for Learning Fractional-Order Nonlinear Dynamical Systems
by: Yaghooti, Bahram, et al.
Published: (2025)