Bayesian Learning of Optimal Policies in Markov Decision Processes with Countably Infinite State-Space
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Adler, Saghar, Subramanian, Vijay |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning to Admit Optimally in an $M/M/k/k+N$ Queueing System with Unknown Service Rate
von: Adler, Saghar, et al.
Veröffentlicht: (2022)
von: Adler, Saghar, et al.
Veröffentlicht: (2022)
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
von: Angelotti, Giorgio, et al.
Veröffentlicht: (2021)
von: Angelotti, Giorgio, et al.
Veröffentlicht: (2021)
On the Global Convergence of Policy Gradient in Average Reward Markov Decision Processes
von: Kumar, Navdeep, et al.
Veröffentlicht: (2024)
von: Kumar, Navdeep, et al.
Veröffentlicht: (2024)
Conformal Off-Policy Evaluation in Markov Decision Processes
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
Bayesian Ambiguity Contraction-based Adaptive Robust Markov Decision Processes for Adversarial Surveillance Missions
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
von: Leon, Vincent, et al.
Veröffentlicht: (2023)
von: Leon, Vincent, et al.
Veröffentlicht: (2023)
Concentration of Cumulative Reward in Markov Decision Processes
von: Sayedana, Borna, et al.
Veröffentlicht: (2024)
von: Sayedana, Borna, et al.
Veröffentlicht: (2024)
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
von: Murthy, Yashaswini, et al.
Veröffentlicht: (2023)
von: Murthy, Yashaswini, et al.
Veröffentlicht: (2023)
Finite Memory Belief Approximation for Optimal Control in Partially Observable Markov Decision Processes
von: Kim, Mintae
Veröffentlicht: (2026)
von: Kim, Mintae
Veröffentlicht: (2026)
Approximate Linear Programming for Decentralized Policy Iteration in Cooperative Multi-agent Markov Decision Processes
von: Mandal, Lakshmi, et al.
Veröffentlicht: (2023)
von: Mandal, Lakshmi, et al.
Veröffentlicht: (2023)
OCMDP: Observation-Constrained Markov Decision Process
von: Wang, Taiyi, et al.
Veröffentlicht: (2024)
von: Wang, Taiyi, et al.
Veröffentlicht: (2024)
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
von: Azeem, Muqsit, et al.
Veröffentlicht: (2024)
von: Azeem, Muqsit, et al.
Veröffentlicht: (2024)
A Physics-Informed Learning Framework to Solve the Infinite-Horizon Optimal Control Problem
von: Fotiadis, Filippos, et al.
Veröffentlicht: (2025)
von: Fotiadis, Filippos, et al.
Veröffentlicht: (2025)
Recursive Gaussian Process State Space Model
von: Zheng, Tengjie, et al.
Veröffentlicht: (2024)
von: Zheng, Tengjie, et al.
Veröffentlicht: (2024)
Learning Markov Processes as Sum-of-Square Forms for Analytical Belief Propagation
von: Amorese, Peter, et al.
Veröffentlicht: (2026)
von: Amorese, Peter, et al.
Veröffentlicht: (2026)
Optimal Bayesian Affine Estimator and Active Learning for the Wiener Model
von: Vakili, Sasan, et al.
Veröffentlicht: (2025)
von: Vakili, Sasan, et al.
Veröffentlicht: (2025)
End-to-End Learning Framework for Solving Non-Markovian Optimal Control
von: Zhang, Xiaole, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaole, et al.
Veröffentlicht: (2025)
Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form
von: Kitamura, Toshinori, et al.
Veröffentlicht: (2024)
von: Kitamura, Toshinori, et al.
Veröffentlicht: (2024)
Compression Method for Deep Diagonal State Space Model Based on $H^2$ Optimal Reduction
von: Sakamoto, Hiroki, et al.
Veröffentlicht: (2025)
von: Sakamoto, Hiroki, et al.
Veröffentlicht: (2025)
Performance of NPG in Countable State-Space Average-Cost RL
von: Murthy, Yashaswini, et al.
Veröffentlicht: (2024)
von: Murthy, Yashaswini, et al.
Veröffentlicht: (2024)
Harnessing Data from Clustered LQR Systems: Personalized and Collaborative Policy Optimization
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2025)
von: Kanakeri, Vinay, et al.
Veröffentlicht: (2025)
LEDRO: LLM-Enhanced Design Space Reduction and Optimization for Analog Circuits
von: Kochar, Dimple Vijay, et al.
Veröffentlicht: (2024)
von: Kochar, Dimple Vijay, et al.
Veröffentlicht: (2024)
Learning Surrogate LPV State-Space Models with Uncertainty Quantification
von: Olucha, E. Javier, et al.
Veröffentlicht: (2026)
von: Olucha, E. Javier, et al.
Veröffentlicht: (2026)
Sequential Bayesian Optimal Experimental Design in Infinite Dimensions via Policy Gradient Reinforcement Learning
von: Shen, Kaichen, et al.
Veröffentlicht: (2026)
von: Shen, Kaichen, et al.
Veröffentlicht: (2026)
Optimal Sample Complexity for Average Reward Markov Decision Processes
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Flipping-based Policy for Chance-Constrained Markov Decision Processes
von: Shen, Xun, et al.
Veröffentlicht: (2024)
von: Shen, Xun, et al.
Veröffentlicht: (2024)
Learning-Based Optimal Control with Performance Guarantees for Unknown Systems with Latent States
von: Lefringhausen, Robert, et al.
Veröffentlicht: (2023)
von: Lefringhausen, Robert, et al.
Veröffentlicht: (2023)
An Optimal Policy for Learning Controllable Dynamics by Exploration
von: Loxley, Peter N.
Veröffentlicht: (2025)
von: Loxley, Peter N.
Veröffentlicht: (2025)
Learning Stable and Robust Linear Parameter-Varying State-Space Models
von: Verhoek, Chris, et al.
Veröffentlicht: (2023)
von: Verhoek, Chris, et al.
Veröffentlicht: (2023)
GenSafe: A Generalizable Safety Enhancer for Safe Reinforcement Learning Algorithms Based on Reduced Order Markov Decision Process Model
von: Zhou, Zhehua, et al.
Veröffentlicht: (2024)
von: Zhou, Zhehua, et al.
Veröffentlicht: (2024)
On Policy Stochasticity in Mutual Information Optimal Control of Linear Systems
von: Enami, Shoju, et al.
Veröffentlicht: (2025)
von: Enami, Shoju, et al.
Veröffentlicht: (2025)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
von: Ibrahim, Sinan, et al.
Veröffentlicht: (2026)
von: Ibrahim, Sinan, et al.
Veröffentlicht: (2026)
Model-Free Learning for the Linear Quadratic Regulator over Rate-Limited Channels
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
von: Ye, Lintao, et al.
Veröffentlicht: (2024)
Federated Causal Representation Learning in State-Space Systems for Decentralized Counterfactual Reasoning
von: Mohamed, Nazal, et al.
Veröffentlicht: (2026)
von: Mohamed, Nazal, et al.
Veröffentlicht: (2026)
On the Relation of State Space Models and Hidden Markov Models
von: Ghojogh, Aydin, et al.
Veröffentlicht: (2026)
von: Ghojogh, Aydin, et al.
Veröffentlicht: (2026)
Layer-Adaptive State Pruning for Deep State Space Models
von: Gwak, Minseon, et al.
Veröffentlicht: (2024)
von: Gwak, Minseon, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning Behavioral Mode Switching Using Optimal Control Based on a Latent Space Objective
von: Remman, Sindre Benjamin, et al.
Veröffentlicht: (2024)
von: Remman, Sindre Benjamin, et al.
Veröffentlicht: (2024)
Space-Filling Regularization for Robust and Interpretable Nonlinear State Space Models
von: Klein, Hermann, et al.
Veröffentlicht: (2025)
von: Klein, Hermann, et al.
Veröffentlicht: (2025)
Improved Monte Carlo Planning via Causal Disentanglement for Structurally-Decomposed Markov Decision Processes
von: Liu, Larkin, et al.
Veröffentlicht: (2024)
von: Liu, Larkin, et al.
Veröffentlicht: (2024)
Optimal Parameter Adaptation for Safety-Critical Control via Safe Barrier Bayesian Optimization
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning to Admit Optimally in an $M/M/k/k+N$ Queueing System with Unknown Service Rate
von: Adler, Saghar, et al.
Veröffentlicht: (2022) -
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
von: Angelotti, Giorgio, et al.
Veröffentlicht: (2021) -
On the Global Convergence of Policy Gradient in Average Reward Markov Decision Processes
von: Kumar, Navdeep, et al.
Veröffentlicht: (2024) -
Conformal Off-Policy Evaluation in Markov Decision Processes
von: Foffano, Daniele, et al.
Veröffentlicht: (2023) -
Bayesian Ambiguity Contraction-based Adaptive Robust Markov Decision Processes for Adversarial Surveillance Missions
von: Choi, Jimin, et al.
Veröffentlicht: (2025)