An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
Fuente:
arXiv
Saved in:
| Main Authors: | Misra, Rahul, Bujorianu, Manuela L., Wisniewski, Rafał |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
by: Mazumdar, Abhijit, et al.
Published: (2024)
by: Mazumdar, Abhijit, et al.
Published: (2024)
Data-Driven Robust Safety Verification for Markov Decision Processes
by: Mazumdar, Abhijit, et al.
Published: (2025)
by: Mazumdar, Abhijit, et al.
Published: (2025)
Markov Decision Process and Approximate Dynamic Programming for a Patient Assignment Scheduling problem
by: O'Reilly, Malgorzata M., et al.
Published: (2024)
by: O'Reilly, Malgorzata M., et al.
Published: (2024)
Robust Correlated Equilibrium: Definition and Computation
by: Misra, Rahul, et al.
Published: (2023)
by: Misra, Rahul, et al.
Published: (2023)
Beyond Average Return in Markov Decision Processes
by: Marthe, Alexandre, et al.
Published: (2023)
by: Marthe, Alexandre, et al.
Published: (2023)
Large Deviations in Safety-Critical Systems with Probabilistic Initial Conditions
by: Gomez, Aitor R., et al.
Published: (2024)
by: Gomez, Aitor R., et al.
Published: (2024)
Optimal Risk-Sensitive Scheduling Policies for Remote Estimation of Autoregressive Markov Processes
by: Dutta, Manali, et al.
Published: (2024)
by: Dutta, Manali, et al.
Published: (2024)
Reoptimization Nearly Solves Weakly Coupled Markov Decision Processes
by: Gast, Nicolas, et al.
Published: (2022)
by: Gast, Nicolas, et al.
Published: (2022)
Stability of Polling Systems for a Large Class of Markovian Switching Policies
by: Avrachenkov, Konstantin, et al.
Published: (2025)
by: Avrachenkov, Konstantin, et al.
Published: (2025)
On Pioneering Works of Albert Shiryaev on Markov Decision Processes and Some Later Developments
by: Feinberg, Eugene A.
Published: (2025)
by: Feinberg, Eugene A.
Published: (2025)
On the Emergence of Linear Behavior in Large-Scale Dynamical Systems via Spatial Averaging
by: Ahmed, Sabbir, et al.
Published: (2025)
by: Ahmed, Sabbir, et al.
Published: (2025)
Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning
by: Srikant, R.
Published: (2024)
by: Srikant, R.
Published: (2024)
Markov Decision Processing Networks
by: Bhambay, Sanidhay, et al.
Published: (2025)
by: Bhambay, Sanidhay, et al.
Published: (2025)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
by: Leon, Vincent, et al.
Published: (2023)
by: Leon, Vincent, et al.
Published: (2023)
Non-Exchangeable Mean Field Markov Decision Processes with common noise : from Bellman equation to quantitative propagation of chaos
by: Mekkaoui, Samy, et al.
Published: (2026)
by: Mekkaoui, Samy, et al.
Published: (2026)
Convex Approximations of Random Constrained Markov Decision Processes
by: Varagapriya, V, et al.
Published: (2025)
by: Varagapriya, V, et al.
Published: (2025)
Operator Splitting for Convex Constrained Markov Decision Processes
by: Grontas, Panagiotis D., et al.
Published: (2024)
by: Grontas, Panagiotis D., et al.
Published: (2024)
Markov Decision Process Design: A Framework for Integrating Strategic and Operational Decisions
by: Brown, Seth, et al.
Published: (2023)
by: Brown, Seth, et al.
Published: (2023)
Quantum Markov Decision Processes: General Theory, Approximations, and Classes of Policies
by: Saldi, Naci, et al.
Published: (2024)
by: Saldi, Naci, et al.
Published: (2024)
A Feedback Control Framework for Incentivised Suburban Parking Utilisation and Urban Core Traffic Relief
by: Satti, Abdul Baseer, et al.
Published: (2025)
by: Satti, Abdul Baseer, et al.
Published: (2025)
Throughput Maximizing Takeoff Scheduling for eVTOL Vehicles in On-Demand Urban Air Mobility Systems
by: Pooladsanj, Milad, et al.
Published: (2025)
by: Pooladsanj, Milad, et al.
Published: (2025)
Trajectory-Based Optimization for Air Traffic Control in the Terminal Maneuvering Area
by: Pang, Yutian, et al.
Published: (2026)
by: Pang, Yutian, et al.
Published: (2026)
Pathwise Relaxed Optimal Control of Rough Differential Equations
by: Chakraborty, Prakash, et al.
Published: (2024)
by: Chakraborty, Prakash, et al.
Published: (2024)
Trajectories and Platoon-forming Algorithm for Intersections with Heterogeneous Autonomous Traffic
by: Joshi, P. C., et al.
Published: (2023)
by: Joshi, P. C., et al.
Published: (2023)
Control policies for a two-stage queueing system with parallel and single server options
by: Lu, Shuwen, et al.
Published: (2026)
by: Lu, Shuwen, et al.
Published: (2026)
Explicit Steady-State Approximations for Parallel Server Systems with Heterogeneous Servers
by: Xu, Yaosheng
Published: (2024)
by: Xu, Yaosheng
Published: (2024)
Balancing Independent and Collaborative Service
by: Lu, Shuwen, et al.
Published: (2026)
by: Lu, Shuwen, et al.
Published: (2026)
The Variational Approach in Filtering and Correlated Noise
by: Srinivasan, Sharan, et al.
Published: (2026)
by: Srinivasan, Sharan, et al.
Published: (2026)
Environmental management and restoration under unified risk and uncertainty using robustified dynamic Orlicz risk
by: Yoshioka, Hidekazu, et al.
Published: (2023)
by: Yoshioka, Hidekazu, et al.
Published: (2023)
A Unified Control Theory Derivation of Discrete-Time Linear Ensemble Kalman Filters
by: Kim, Jin Won
Published: (2026)
by: Kim, Jin Won
Published: (2026)
Revisiting Stochastic Realization Theory using Functional Itô Calculus
by: Veeravalli, Tanya, et al.
Published: (2024)
by: Veeravalli, Tanya, et al.
Published: (2024)
SIS epidemics on open networks: A replacement-based approximation
by: Vizuete, Renato, et al.
Published: (2024)
by: Vizuete, Renato, et al.
Published: (2024)
Distributionally Robust Safety Verification for Markov Decision Processes
by: Mazumdar, Abhijit, et al.
Published: (2024)
by: Mazumdar, Abhijit, et al.
Published: (2024)
Bounding the Difference between the Values of Robust and Non-Robust Markov Decision Problems
by: Neufeld, Ariel, et al.
Published: (2023)
by: Neufeld, Ariel, et al.
Published: (2023)
Decentralized State-Dependent Markov Chain Synthesis with an Application to Swarm Guidance
by: Uzun, Samet, et al.
Published: (2020)
by: Uzun, Samet, et al.
Published: (2020)
Linear Algebraic Truncation Algorithm with A Posteriori Error Bounds for Computing Markov Chain Equilibrium Gradients
by: Mahdian, Saied, et al.
Published: (2025)
by: Mahdian, Saied, et al.
Published: (2025)
A Markov Decision Process Framework for Enhancing Power System Resilience during Wildfires under Decision-Dependent Uncertainty
by: Zhao, Xinyi, et al.
Published: (2026)
by: Zhao, Xinyi, et al.
Published: (2026)
Markov control of continuous time Markov processes with long run functionals by time discretization
by: Stettner, Lukasz
Published: (2025)
by: Stettner, Lukasz
Published: (2025)
Almost Sure Convergence of Networked Policy Gradient over Time-Varying Networks in Markov Potential Games
by: Aydin, Sarper, et al.
Published: (2024)
by: Aydin, Sarper, et al.
Published: (2024)
Derivative Estimation from Coarse, Irregular, Noisy Samples: An MLE-Spline Approach
by: Avrachenkov, Konstantin E., et al.
Published: (2025)
by: Avrachenkov, Konstantin E., et al.
Published: (2025)
Similar Items
-
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
by: Mazumdar, Abhijit, et al.
Published: (2024) -
Data-Driven Robust Safety Verification for Markov Decision Processes
by: Mazumdar, Abhijit, et al.
Published: (2025) -
Markov Decision Process and Approximate Dynamic Programming for a Patient Assignment Scheduling problem
by: O'Reilly, Malgorzata M., et al.
Published: (2024) -
Robust Correlated Equilibrium: Definition and Computation
by: Misra, Rahul, et al.
Published: (2023) -
Beyond Average Return in Markov Decision Processes
by: Marthe, Alexandre, et al.
Published: (2023)