Model-Based Exploration in Monitored Markov Decision Processes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kazemipour, Alireza, Parisi, Simone, Taylor, Matthew E., Bowling, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Monitored Markov Decision Processes
von: Parisi, Simone, et al.
Veröffentlicht: (2024)
von: Parisi, Simone, et al.
Veröffentlicht: (2024)
Beyond Optimism: Exploration With Partially Observable Rewards
von: Parisi, Simone, et al.
Veröffentlicht: (2024)
von: Parisi, Simone, et al.
Veröffentlicht: (2024)
Generalization in Monitored Markov Decision Processes (Mon-MDPs)
von: Mohammedalamen, Montaser, et al.
Veröffentlicht: (2025)
von: Mohammedalamen, Montaser, et al.
Veröffentlicht: (2025)
Learn A Flexible Exploration Model for Parameterized Action Markov Decision Processes
von: Wang, Zijian, et al.
Veröffentlicht: (2025)
von: Wang, Zijian, et al.
Veröffentlicht: (2025)
Logarithmic Regret of Exploration in Average Reward Markov Decision Processes
von: Boone, Victor, et al.
Veröffentlicht: (2025)
von: Boone, Victor, et al.
Veröffentlicht: (2025)
Geometric Active Exploration in Markov Decision Processes: the Benefit of Abstraction
von: De Santi, Riccardo, et al.
Veröffentlicht: (2024)
von: De Santi, Riccardo, et al.
Veröffentlicht: (2024)
Generalized Linear Markov Decision Process
von: Zhang, Sinian, et al.
Veröffentlicht: (2025)
von: Zhang, Sinian, et al.
Veröffentlicht: (2025)
Federated Control in Markov Decision Processes
von: Jin, Hao, et al.
Veröffentlicht: (2024)
von: Jin, Hao, et al.
Veröffentlicht: (2024)
Concentration of Cumulative Reward in Markov Decision Processes
von: Sayedana, Borna, et al.
Veröffentlicht: (2024)
von: Sayedana, Borna, et al.
Veröffentlicht: (2024)
Learning in Markov Decision Processes with Exogenous Dynamics
von: Maran, Davide, et al.
Veröffentlicht: (2026)
von: Maran, Davide, et al.
Veröffentlicht: (2026)
Increasing Information for Model Predictive Control with Semi-Markov Decision Processes
von: Hosseinkhan-Boucher, Rémy, et al.
Veröffentlicht: (2025)
von: Hosseinkhan-Boucher, Rémy, et al.
Veröffentlicht: (2025)
Optimal Decision Tree Policies for Markov Decision Processes
von: Vos, Daniël, et al.
Veröffentlicht: (2023)
von: Vos, Daniël, et al.
Veröffentlicht: (2023)
Policy Testing in Markov Decision Processes
von: Ariu, Kaito, et al.
Veröffentlicht: (2025)
von: Ariu, Kaito, et al.
Veröffentlicht: (2025)
Markov Decision Processes under External Temporal Processes
von: Ayyagari, Ranga Shaarad, et al.
Veröffentlicht: (2023)
von: Ayyagari, Ranga Shaarad, et al.
Veröffentlicht: (2023)
The regret lower bound for communicating Markov Decision Processes
von: Boone, Victor, et al.
Veröffentlicht: (2025)
von: Boone, Victor, et al.
Veröffentlicht: (2025)
An Orthogonal Learner for Individualized Outcomes in Markov Decision Processes
von: Javurek, Emil, et al.
Veröffentlicht: (2025)
von: Javurek, Emil, et al.
Veröffentlicht: (2025)
Initial Distribution Sensitivity of Constrained Markov Decision Processes
von: Tercan, Alperen, et al.
Veröffentlicht: (2025)
von: Tercan, Alperen, et al.
Veröffentlicht: (2025)
Improving Controller Generalization with Dimensionless Markov Decision Processes
von: Charvet, Valentin, et al.
Veröffentlicht: (2025)
von: Charvet, Valentin, et al.
Veröffentlicht: (2025)
Horizon-Free Regret for Linear Markov Decision Processes
von: Zhang, Zihan, et al.
Veröffentlicht: (2024)
von: Zhang, Zihan, et al.
Veröffentlicht: (2024)
Learning Utilities from Demonstrations in Markov Decision Processes
von: Lazzati, Filippo, et al.
Veröffentlicht: (2024)
von: Lazzati, Filippo, et al.
Veröffentlicht: (2024)
Achieving Constant Regret in Linear Markov Decision Processes
von: Zhang, Weitong, et al.
Veröffentlicht: (2024)
von: Zhang, Weitong, et al.
Veröffentlicht: (2024)
Policy Gradient for Robust Markov Decision Processes
von: Wang, Qiuhao, et al.
Veröffentlicht: (2024)
von: Wang, Qiuhao, et al.
Veröffentlicht: (2024)
No-Regret Thompson Sampling for Finite-Horizon Markov Decision Processes with Gaussian Processes
von: Bayrooti, Jasmine, et al.
Veröffentlicht: (2025)
von: Bayrooti, Jasmine, et al.
Veröffentlicht: (2025)
A Markov Decision Process for Variable Selection in Branch & Bound
von: Strang, Paul, et al.
Veröffentlicht: (2025)
von: Strang, Paul, et al.
Veröffentlicht: (2025)
Transition Transfer $Q$-Learning for Composite Markov Decision Processes
von: Chai, Jinhang, et al.
Veröffentlicht: (2025)
von: Chai, Jinhang, et al.
Veröffentlicht: (2025)
Best-of-Both-Worlds for Heavy-Tailed Markov Decision Processes
von: Chen, Yu, et al.
Veröffentlicht: (2026)
von: Chen, Yu, et al.
Veröffentlicht: (2026)
Performance Improvement Bounds for Lipschitz Configurable Markov Decision Processes
von: Metelli, Alberto Maria
Veröffentlicht: (2024)
von: Metelli, Alberto Maria
Veröffentlicht: (2024)
Transition Constrained Bayesian Optimization via Markov Decision Processes
von: Folch, Jose Pablo, et al.
Veröffentlicht: (2024)
von: Folch, Jose Pablo, et al.
Veröffentlicht: (2024)
Learning Markov Decision Processes under Fully Bandit Feedback
von: Zhuo, Zhengjia, et al.
Veröffentlicht: (2026)
von: Zhuo, Zhengjia, et al.
Veröffentlicht: (2026)
Fair Resource Allocation in Weakly Coupled Markov Decision Processes
von: Tu, Xiaohui, et al.
Veröffentlicht: (2024)
von: Tu, Xiaohui, et al.
Veröffentlicht: (2024)
Rate-Optimal Policy Optimization for Linear Markov Decision Processes
von: Sherman, Uri, et al.
Veröffentlicht: (2023)
von: Sherman, Uri, et al.
Veröffentlicht: (2023)
Laplacian Representations for Decision-Time Planning
von: Shehmar, Dikshant, et al.
Veröffentlicht: (2026)
von: Shehmar, Dikshant, et al.
Veröffentlicht: (2026)
Weakly Time-Coupled Approximation of Markov Decision Processes
von: Soheili, Negar, et al.
Veröffentlicht: (2026)
von: Soheili, Negar, et al.
Veröffentlicht: (2026)
Online Markov Decision Processes with Terminal Law Constraints
von: Moreno, Bianca Marin, et al.
Veröffentlicht: (2026)
von: Moreno, Bianca Marin, et al.
Veröffentlicht: (2026)
A Theoretical Analysis of State Similarity Between Markov Decision Processes
von: Tao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Tao, Zhenyu, et al.
Veröffentlicht: (2025)
Learning Deterministic Policies with Policy Gradients in Constrained Markov Decision Processes
von: Montenegro, Alessandro, et al.
Veröffentlicht: (2025)
von: Montenegro, Alessandro, et al.
Veröffentlicht: (2025)
Optimistic Actor-Critic with Parametric Policies for Linear Markov Decision Processes
von: Lin, Max Qiushi, et al.
Veröffentlicht: (2026)
von: Lin, Max Qiushi, et al.
Veröffentlicht: (2026)
Interaction-Grounded Learning for Contextual Markov Decision Processes with Personalized Feedback
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2026)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2026)
Impact of Markov Decision Process Design on Sim-to-Real Reinforcement Learning
von: Krau, Tatjana, et al.
Veröffentlicht: (2026)
von: Krau, Tatjana, et al.
Veröffentlicht: (2026)
Topology-Aware State Abstraction with Tangle Cores for Markov Decision Processes
von: Shihab, Ibne Farabi, et al.
Veröffentlicht: (2026)
von: Shihab, Ibne Farabi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Monitored Markov Decision Processes
von: Parisi, Simone, et al.
Veröffentlicht: (2024) -
Beyond Optimism: Exploration With Partially Observable Rewards
von: Parisi, Simone, et al.
Veröffentlicht: (2024) -
Generalization in Monitored Markov Decision Processes (Mon-MDPs)
von: Mohammedalamen, Montaser, et al.
Veröffentlicht: (2025) -
Learn A Flexible Exploration Model for Parameterized Action Markov Decision Processes
von: Wang, Zijian, et al.
Veröffentlicht: (2025) -
Logarithmic Regret of Exploration in Average Reward Markov Decision Processes
von: Boone, Victor, et al.
Veröffentlicht: (2025)