Saved in:
| Main Authors: | Santos, Pedro P., Sardinha, Alberto, Melo, Francisco S. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2409.15128 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Solving General-Utility Markov Decision Processes in the Single-Trial Regime with Online Planning
by: Santos, Pedro P., et al.
Published: (2025)
by: Santos, Pedro P., et al.
Published: (2025)
Entropic Risk-Aware Monte Carlo Tree Search
by: Santos, Pedro P., et al.
Published: (2026)
by: Santos, Pedro P., et al.
Published: (2026)
Learning Utilities from Demonstrations in Markov Decision Processes
by: Lazzati, Filippo, et al.
Published: (2024)
by: Lazzati, Filippo, et al.
Published: (2024)
Implicit Repair with Reinforcement Learning in Emergent Communication
by: Vital, Fábio, et al.
Published: (2025)
by: Vital, Fábio, et al.
Published: (2025)
Quantum Speedups in Regret Analysis of Infinite Horizon Average-Reward Markov Decision Processes
by: Ganguly, Bhargav, et al.
Published: (2023)
by: Ganguly, Bhargav, et al.
Published: (2023)
Regret Analysis of Policy Gradient Algorithm for Infinite Horizon Average Reward Markov Decision Processes
by: Bai, Qinbo, et al.
Published: (2023)
by: Bai, Qinbo, et al.
Published: (2023)
Horizon-Free Regret for Linear Markov Decision Processes
by: Zhang, Zihan, et al.
Published: (2024)
by: Zhang, Zihan, et al.
Published: (2024)
Distributed Value Decomposition Networks with Networked Agents
by: Varela, Guilherme S., et al.
Published: (2025)
by: Varela, Guilherme S., et al.
Published: (2025)
Networked Agents in the Dark: Team Value Learning under Partial Observability
by: Varela, Guilherme S., et al.
Published: (2025)
by: Varela, Guilherme S., et al.
Published: (2025)
Improved Sample Complexity Analysis of Natural Policy Gradient Algorithm with General Parameterization for Infinite Horizon Discounted Reward Markov Decision Processes
by: Mondal, Washim Uddin, et al.
Published: (2023)
by: Mondal, Washim Uddin, et al.
Published: (2023)
Risk-sensitive Markov Decision Process and Learning under General Utility Functions
by: Wu, Zhengqi, et al.
Published: (2023)
by: Wu, Zhengqi, et al.
Published: (2023)
No-Regret Thompson Sampling for Finite-Horizon Markov Decision Processes with Gaussian Processes
by: Bayrooti, Jasmine, et al.
Published: (2025)
by: Bayrooti, Jasmine, et al.
Published: (2025)
A policy gradient approach for Finite Horizon Constrained Markov Decision Processes
by: Guin, Soumyajit, et al.
Published: (2022)
by: Guin, Soumyajit, et al.
Published: (2022)
Generalized Linear Markov Decision Process
by: Zhang, Sinian, et al.
Published: (2025)
by: Zhang, Sinian, et al.
Published: (2025)
Thompson Sampling for Infinite-Horizon Discounted Decision Processes
by: Adelman, Daniel, et al.
Published: (2024)
by: Adelman, Daniel, et al.
Published: (2024)
Performance Improvement Bounds for Lipschitz Configurable Markov Decision Processes
by: Metelli, Alberto Maria
Published: (2024)
by: Metelli, Alberto Maria
Published: (2024)
Bayesian Learning of Optimal Policies in Markov Decision Processes with Countably Infinite State-Space
by: Adler, Saghar, et al.
Published: (2023)
by: Adler, Saghar, et al.
Published: (2023)
RecBayes: Recurrent Bayesian Ad Hoc Teamwork in Large Partially Observable Domains
by: Ribeiro, João G., et al.
Published: (2025)
by: Ribeiro, João G., et al.
Published: (2025)
Improving Controller Generalization with Dimensionless Markov Decision Processes
by: Charvet, Valentin, et al.
Published: (2025)
by: Charvet, Valentin, et al.
Published: (2025)
Monitored Markov Decision Processes
by: Parisi, Simone, et al.
Published: (2024)
by: Parisi, Simone, et al.
Published: (2024)
SPOT: Scalable Policy Optimization with Trees for Markov Decision Processes
by: Xiong, Xuyuan, et al.
Published: (2025)
by: Xiong, Xuyuan, et al.
Published: (2025)
Learning Constrained Markov Decision Processes With Non-stationary Rewards and Constraints
by: Stradi, Francesco Emanuele, et al.
Published: (2024)
by: Stradi, Francesco Emanuele, et al.
Published: (2024)
Learning Deterministic Policies with Policy Gradients in Constrained Markov Decision Processes
by: Montenegro, Alessandro, et al.
Published: (2025)
by: Montenegro, Alessandro, et al.
Published: (2025)
Federated Control in Markov Decision Processes
by: Jin, Hao, et al.
Published: (2024)
by: Jin, Hao, et al.
Published: (2024)
Solving Robust Markov Decision Processes: Generic, Reliable, Efficient
by: Meggendorfer, Tobias, et al.
Published: (2024)
by: Meggendorfer, Tobias, et al.
Published: (2024)
Learning in Markov Decision Processes with Exogenous Dynamics
by: Maran, Davide, et al.
Published: (2026)
by: Maran, Davide, et al.
Published: (2026)
Optimal Decision Tree Policies for Markov Decision Processes
by: Vos, Daniël, et al.
Published: (2023)
by: Vos, Daniël, et al.
Published: (2023)
Policy Testing in Markov Decision Processes
by: Ariu, Kaito, et al.
Published: (2025)
by: Ariu, Kaito, et al.
Published: (2025)
Markov Decision Processes under External Temporal Processes
by: Ayyagari, Ranga Shaarad, et al.
Published: (2023)
by: Ayyagari, Ranga Shaarad, et al.
Published: (2023)
Quantum Logic Gate Synthesis as a Markov Decision Process
by: Alam, M. Sohaib, et al.
Published: (2019)
by: Alam, M. Sohaib, et al.
Published: (2019)
Achieving Constant Regret in Linear Markov Decision Processes
by: Zhang, Weitong, et al.
Published: (2024)
by: Zhang, Weitong, et al.
Published: (2024)
The regret lower bound for communicating Markov Decision Processes
by: Boone, Victor, et al.
Published: (2025)
by: Boone, Victor, et al.
Published: (2025)
An Orthogonal Learner for Individualized Outcomes in Markov Decision Processes
by: Javurek, Emil, et al.
Published: (2025)
by: Javurek, Emil, et al.
Published: (2025)
Initial Distribution Sensitivity of Constrained Markov Decision Processes
by: Tercan, Alperen, et al.
Published: (2025)
by: Tercan, Alperen, et al.
Published: (2025)
Model-Based Exploration in Monitored Markov Decision Processes
by: Kazemipour, Alireza, et al.
Published: (2025)
by: Kazemipour, Alireza, et al.
Published: (2025)
Concentration of Cumulative Reward in Markov Decision Processes
by: Sayedana, Borna, et al.
Published: (2024)
by: Sayedana, Borna, et al.
Published: (2024)
Policy Gradient for Robust Markov Decision Processes
by: Wang, Qiuhao, et al.
Published: (2024)
by: Wang, Qiuhao, et al.
Published: (2024)
Transition Constrained Bayesian Optimization via Markov Decision Processes
by: Folch, Jose Pablo, et al.
Published: (2024)
by: Folch, Jose Pablo, et al.
Published: (2024)
Fair Resource Allocation in Weakly Coupled Markov Decision Processes
by: Tu, Xiaohui, et al.
Published: (2024)
by: Tu, Xiaohui, et al.
Published: (2024)
Geometric Active Exploration in Markov Decision Processes: the Benefit of Abstraction
by: De Santi, Riccardo, et al.
Published: (2024)
by: De Santi, Riccardo, et al.
Published: (2024)
Similar Items
-
Solving General-Utility Markov Decision Processes in the Single-Trial Regime with Online Planning
by: Santos, Pedro P., et al.
Published: (2025) -
Entropic Risk-Aware Monte Carlo Tree Search
by: Santos, Pedro P., et al.
Published: (2026) -
Learning Utilities from Demonstrations in Markov Decision Processes
by: Lazzati, Filippo, et al.
Published: (2024) -
Implicit Repair with Reinforcement Learning in Emergent Communication
by: Vital, Fábio, et al.
Published: (2025) -
Quantum Speedups in Regret Analysis of Infinite Horizon Average-Reward Markov Decision Processes
by: Ganguly, Bhargav, et al.
Published: (2023)