An Adaptive Method for Contextual Stochastic Multi-armed Bandits with Rewards Generated by a Linear Dynamical System
Fuente:
arXiv
Guardado en:
| Autores principales: | Gornet, Jonathan, Hosseinzadeh, Mehdi, Sinopoli, Bruno |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Restless Bandit Problem with Rewards Generated by a Linear Gaussian Dynamical System
por: Gornet, Jonathan, et al.
Publicado: (2024)
por: Gornet, Jonathan, et al.
Publicado: (2024)
An Exploration-free Method for a Linear Stochastic Bandit Driven by a Linear Gaussian Dynamical System
por: Gornet, Jonathan, et al.
Publicado: (2025)
por: Gornet, Jonathan, et al.
Publicado: (2025)
A Control Theory inspired Exploration Method for a Linear Bandit driven by a Linear Gaussian Dynamical System
por: Gornet, Jonathan, et al.
Publicado: (2025)
por: Gornet, Jonathan, et al.
Publicado: (2025)
HyperController: A Hyperparameter Controller for Fast and Stable Training of Reinforcement Learning Neural Networks
por: Gornet, Jonathan, et al.
Publicado: (2025)
por: Gornet, Jonathan, et al.
Publicado: (2025)
Switched Linear Ensemble Systems and Structural Controllability
por: Yin, Haoyu, et al.
Publicado: (2025)
por: Yin, Haoyu, et al.
Publicado: (2025)
Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
por: Manupriya, Piyushi, et al.
Publicado: (2025)
por: Manupriya, Piyushi, et al.
Publicado: (2025)
On Zero-sum Game Representation for Replicator Dynamics
por: Yin, Haoyu, et al.
Publicado: (2025)
por: Yin, Haoyu, et al.
Publicado: (2025)
On Permanence of Conservative Replicator Dynamics with Four Strategies
por: Yin, Haoyu, et al.
Publicado: (2026)
por: Yin, Haoyu, et al.
Publicado: (2026)
A Data-Integrated Framework for Learning Fractional-Order Nonlinear Dynamical Systems
por: Yaghooti, Bahram, et al.
Publicado: (2025)
por: Yaghooti, Bahram, et al.
Publicado: (2025)
Learning to Sparsify Stochastic Linear Bandits
por: Wang, Zhengmiao, et al.
Publicado: (2026)
por: Wang, Zhengmiao, et al.
Publicado: (2026)
Model-Free Learning and Optimal Policy Design in Multi-Agent MDPs Under Probabilistic Agent Dropout
por: Fiscko, Carmel, et al.
Publicado: (2023)
por: Fiscko, Carmel, et al.
Publicado: (2023)
Second Moment Polytopic Systems: Generalization of Uncertain Stochastic Linear Dynamics
por: Ito, Yuji, et al.
Publicado: (2022)
por: Ito, Yuji, et al.
Publicado: (2022)
Provably-Stable Neural Network-Based Control of Nonlinear Systems
por: Li, Anran, et al.
Publicado: (2025)
por: Li, Anran, et al.
Publicado: (2025)
Dynamic Regressor Extension and Mixing-based Re-design of Adaptive Observer for Affine Systems
por: Tavan, Mehdi
Publicado: (2025)
por: Tavan, Mehdi
Publicado: (2025)
Discrete-Time Implementation of Explicit Reference Governor
por: Momani, Mu'taz A., et al.
Publicado: (2024)
por: Momani, Mu'taz A., et al.
Publicado: (2024)
Approximate Solution Methods for the Average Reward Criterion in Optimal Tracking Control of Linear Systems
por: Nguyen, Duc Cuong
Publicado: (2025)
por: Nguyen, Duc Cuong
Publicado: (2025)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
por: Choi, Sunmook, et al.
Publicado: (2025)
por: Choi, Sunmook, et al.
Publicado: (2025)
Distance Between Stochastic Linear Systems
por: Renganathan, Venkatraman, et al.
Publicado: (2025)
por: Renganathan, Venkatraman, et al.
Publicado: (2025)
The Stochastic Occupation Kernel Method for System Identification
por: Wells, Michael, et al.
Publicado: (2024)
por: Wells, Michael, et al.
Publicado: (2024)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
por: Adams, Katherine B., et al.
Publicado: (2025)
por: Adams, Katherine B., et al.
Publicado: (2025)
Multi-Agent Stage-wise Conservative Linear Bandits
por: Afsharrad, Amirhossein, et al.
Publicado: (2025)
por: Afsharrad, Amirhossein, et al.
Publicado: (2025)
Multi-Objective Planning with Contextual Lexicographic Reward Preferences
por: Rustagi, Pulkit, et al.
Publicado: (2025)
por: Rustagi, Pulkit, et al.
Publicado: (2025)
Observer-Based Stabilization for Linear Multi-Agent Dynamical Systems Using Generalized Frequency Variables
por: Tran, G. Q. Bao, et al.
Publicado: (2026)
por: Tran, G. Q. Bao, et al.
Publicado: (2026)
A Stochastic Robust Adaptive Systems Level Approach to Stabilizing Large-Scale Uncertain Markovian Jump Linear Systems
por: Han, SooJean, et al.
Publicado: (2024)
por: Han, SooJean, et al.
Publicado: (2024)
System Identification for Continuous-time Linear Dynamical Systems
por: Halmos, Peter, et al.
Publicado: (2023)
por: Halmos, Peter, et al.
Publicado: (2023)
Multi-armed Bandit for Stochastic Shortest Path in Mixed Autonomy
por: Bai, Yu, et al.
Publicado: (2025)
por: Bai, Yu, et al.
Publicado: (2025)
Optimal Covariance Steering of Linear Stochastic Systems with Hybrid Transitions
por: Yu, Hongzhe, et al.
Publicado: (2024)
por: Yu, Hongzhe, et al.
Publicado: (2024)
Scenario Reduction with Guarantees for Stochastic Optimal Control of Linear Systems
por: Cordiano, Francesco, et al.
Publicado: (2024)
por: Cordiano, Francesco, et al.
Publicado: (2024)
Dynamic Connectivity and Local Frequency Strength under Stochastic Variations
por: Pinheiro, Bruno, et al.
Publicado: (2026)
por: Pinheiro, Bruno, et al.
Publicado: (2026)
Distributed Affine Formation Control of Linear Multi-agent Systems with Adaptive Event-triggering
por: Liu, Chenjun, et al.
Publicado: (2025)
por: Liu, Chenjun, et al.
Publicado: (2025)
Robust Model-Free Control Framework with Safety Constraints for a Fully Electric Linear Actuator System
por: Shahna, Mehdi Heydari, et al.
Publicado: (2024)
por: Shahna, Mehdi Heydari, et al.
Publicado: (2024)
Analyzing the Impact of Computation in Adaptive Dynamic Programming for Stochastic LQR Problem
por: Cao, Wenhan, et al.
Publicado: (2024)
por: Cao, Wenhan, et al.
Publicado: (2024)
Reachability and Controllability Analysis of the State Covariance for Linear Stochastic Systems
por: Liu, Fengjiao, et al.
Publicado: (2024)
por: Liu, Fengjiao, et al.
Publicado: (2024)
Optimal Covariance Steering for Discrete-Time Linear Stochastic Systems
por: Liu, Fengjiao, et al.
Publicado: (2022)
por: Liu, Fengjiao, et al.
Publicado: (2022)
Incentive Compatibility in Stochastic Dynamic Systems
por: Ma, Ke, et al.
Publicado: (2019)
por: Ma, Ke, et al.
Publicado: (2019)
Maximizing Reach-Avoid Probabilities for Linear Stochastic Systems via Control Architectures
por: Schmid, Niklas, et al.
Publicado: (2026)
por: Schmid, Niklas, et al.
Publicado: (2026)
Koopman Spectral Analysis and System Identification for Stochastic Dynamical Systems via Yosida Approximation of Generators
por: Zhou, Jun, et al.
Publicado: (2025)
por: Zhou, Jun, et al.
Publicado: (2025)
An LP-based Sampling Policy for Multi-Armed Bandits with Side-Observations and Stochastic Availability
por: Soni, Ashutosh, et al.
Publicado: (2026)
por: Soni, Ashutosh, et al.
Publicado: (2026)
On Reward-Balancing Methods for Reinforcement Learning
por: Baroncini, Simone, et al.
Publicado: (2026)
por: Baroncini, Simone, et al.
Publicado: (2026)
Conformal Prediction-Based MPC for Stochastic Linear Systems
por: Vogel, Lukas, et al.
Publicado: (2025)
por: Vogel, Lukas, et al.
Publicado: (2025)
Ejemplares similares
-
Restless Bandit Problem with Rewards Generated by a Linear Gaussian Dynamical System
por: Gornet, Jonathan, et al.
Publicado: (2024) -
An Exploration-free Method for a Linear Stochastic Bandit Driven by a Linear Gaussian Dynamical System
por: Gornet, Jonathan, et al.
Publicado: (2025) -
A Control Theory inspired Exploration Method for a Linear Bandit driven by a Linear Gaussian Dynamical System
por: Gornet, Jonathan, et al.
Publicado: (2025) -
HyperController: A Hyperparameter Controller for Fast and Stable Training of Reinforcement Learning Neural Networks
por: Gornet, Jonathan, et al.
Publicado: (2025) -
Switched Linear Ensemble Systems and Structural Controllability
por: Yin, Haoyu, et al.
Publicado: (2025)