Off-Policy Temporal Difference Learning for Perturbed Markov Decision Processes: Theoretical Insights and Extensive Simulations
Fuente:
arXiv
Saved in:
| Main Authors: | Forootani, Ali, Iervolino, Raffaele, Tipaldi, Massimo, Khosravi, Mohammad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Asynchronous Federated Learning: A Scalable Approach for Decentralized Machine Learning
by: Forootani, Ali, et al.
Published: (2024)
by: Forootani, Ali, et al.
Published: (2024)
Synthetic Time Series Forecasting with Transformer Architectures: Extensive Simulation Benchmarks
by: Forootani, Ali, et al.
Published: (2025)
by: Forootani, Ali, et al.
Published: (2025)
Learnable Koopman-Enhanced Transformer-Based Time Series Forecasting with Spectral Control
by: Forootani, Ali, et al.
Published: (2026)
by: Forootani, Ali, et al.
Published: (2026)
Asynchronous Federated Learning with non-convex client objective functions and heterogeneous dataset
by: Forootani, Ali, et al.
Published: (2025)
by: Forootani, Ali, et al.
Published: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
by: Foffano, Daniele, et al.
Published: (2023)
by: Foffano, Daniele, et al.
Published: (2023)
Policy Gradient Methods for Information-Theoretic Opacity in Markov Decision Processes
by: Shi, Chongyang, et al.
Published: (2025)
by: Shi, Chongyang, et al.
Published: (2025)
Information-Theoretic Opacity-Enforcement in Markov Decision Processes
by: Shi, Chongyang, et al.
Published: (2024)
by: Shi, Chongyang, et al.
Published: (2024)
Causal Temporal Reasoning for Markov Decision Processes
by: Kazemi, Milad, et al.
Published: (2022)
by: Kazemi, Milad, et al.
Published: (2022)
Learning Robust Policies for Uncertain Parametric Markov Decision Processes
by: Rickard, Luke, et al.
Published: (2023)
by: Rickard, Luke, et al.
Published: (2023)
Differentially Private Reward Functions in Policy Synthesis for Markov Decision Processes
by: Benvenuti, Alexander, et al.
Published: (2023)
by: Benvenuti, Alexander, et al.
Published: (2023)
Bio-Eng-LMM AI Assist chatbot: A Comprehensive Tool for Research and Education
by: Forootani, Ali, et al.
Published: (2024)
by: Forootani, Ali, et al.
Published: (2024)
Approximate Bilevel Difference Convex Programming for Bayesian Risk Markov Decision Processes
by: Lin, Yifan, et al.
Published: (2023)
by: Lin, Yifan, et al.
Published: (2023)
Optimal Control of Markov Decision Processes for Efficiency with Linear Temporal Logic Tasks
by: Chen, Yu, et al.
Published: (2024)
by: Chen, Yu, et al.
Published: (2024)
Quantum Markov Decision Processes: General Theory, Approximations, and Classes of Policies
by: Saldi, Naci, et al.
Published: (2024)
by: Saldi, Naci, et al.
Published: (2024)
RE-LLM: Integrating Large Language Models into Renewable Energy Systems
by: Forootani, Ali, et al.
Published: (2025)
by: Forootani, Ali, et al.
Published: (2025)
Entropy Rate Maximization of Markov Decision Processes under Linear Temporal Logic Tasks
by: Chen, Yu, et al.
Published: (2022)
by: Chen, Yu, et al.
Published: (2022)
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
by: Misra, Rahul, et al.
Published: (2025)
by: Misra, Rahul, et al.
Published: (2025)
Bayesian Learning of Optimal Policies in Markov Decision Processes with Countably Infinite State-Space
by: Adler, Saghar, et al.
Published: (2023)
by: Adler, Saghar, et al.
Published: (2023)
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning
by: Baheri, Ali
Published: (2025)
by: Baheri, Ali
Published: (2025)
On the Global Convergence of Policy Gradient in Average Reward Markov Decision Processes
by: Kumar, Navdeep, et al.
Published: (2024)
by: Kumar, Navdeep, et al.
Published: (2024)
Convex Approximations of Random Constrained Markov Decision Processes
by: Varagapriya, V, et al.
Published: (2025)
by: Varagapriya, V, et al.
Published: (2025)
Operator Splitting for Convex Constrained Markov Decision Processes
by: Grontas, Panagiotis D., et al.
Published: (2024)
by: Grontas, Panagiotis D., et al.
Published: (2024)
Markov Decision Process Design: A Framework for Integrating Strategic and Operational Decisions
by: Brown, Seth, et al.
Published: (2023)
by: Brown, Seth, et al.
Published: (2023)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
by: Leon, Vincent, et al.
Published: (2023)
by: Leon, Vincent, et al.
Published: (2023)
Distributionally Robust Safety Verification for Markov Decision Processes
by: Mazumdar, Abhijit, et al.
Published: (2024)
by: Mazumdar, Abhijit, et al.
Published: (2024)
Real-time Building Energy Storage Scheduling under Electrical Load Uncertainty: A Dynamic Markov Decision Process Approach with Comprehensive Analysis of Different Pricing Policies
by: Sharadga, Hussein, et al.
Published: (2023)
by: Sharadga, Hussein, et al.
Published: (2023)
Intermittently Observable Markov Decision Processes
by: Chen, Gongpu, et al.
Published: (2023)
by: Chen, Gongpu, et al.
Published: (2023)
Data-Driven Robust Safety Verification for Markov Decision Processes
by: Mazumdar, Abhijit, et al.
Published: (2025)
by: Mazumdar, Abhijit, et al.
Published: (2025)
A Markov Decision Process Model for Intrusion Tolerance Problems
by: Kreidl, Patrick
Published: (2025)
by: Kreidl, Patrick
Published: (2025)
Active Inference through Incentive Design in Markov Decision Processes
by: Wei, Xinyi, et al.
Published: (2025)
by: Wei, Xinyi, et al.
Published: (2025)
Accelerating Adaptive Systems via Normalized Parameter Estimation Laws
by: Boveiri, Mohammad, et al.
Published: (2025)
by: Boveiri, Mohammad, et al.
Published: (2025)
Probabilistic Formulations for System Identification of Linear Dynamics with Bilinear Observation Models
by: Liu, Diyou, et al.
Published: (2025)
by: Liu, Diyou, et al.
Published: (2025)
Learning Algorithms for Verification of Markov Decision Processes
by: Brázdil, Tomáš, et al.
Published: (2024)
by: Brázdil, Tomáš, et al.
Published: (2024)
Concentration of Cumulative Reward in Markov Decision Processes
by: Sayedana, Borna, et al.
Published: (2024)
by: Sayedana, Borna, et al.
Published: (2024)
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
by: Angelotti, Giorgio, et al.
Published: (2021)
by: Angelotti, Giorgio, et al.
Published: (2021)
Optimized Task Assignment and Predictive Maintenance for Industrial Machines using Markov Decision Process
by: Nasir, Ali, et al.
Published: (2024)
by: Nasir, Ali, et al.
Published: (2024)
Bridging Impulse Control of Piecewise Deterministic Markov Processes and Markov Decision Processes: Frameworks, Extensions, and Open Challenges
by: Cleynen, Alice, et al.
Published: (2025)
by: Cleynen, Alice, et al.
Published: (2025)
Decentralized Cooperative Localization for Multi-Robot Systems with Asynchronous Sensor Fusion
by: Khosravi, Nivand, et al.
Published: (2026)
by: Khosravi, Nivand, et al.
Published: (2026)
Compositional Planning for Logically Constrained Multi-Agent Markov Decision Processes
by: Kalagarla, Krishna C., et al.
Published: (2024)
by: Kalagarla, Krishna C., et al.
Published: (2024)
Economic Model Predictive Control as a Solution to Markov Decision Processes
by: Reinhardt, Dirk, et al.
Published: (2024)
by: Reinhardt, Dirk, et al.
Published: (2024)
Similar Items
-
Asynchronous Federated Learning: A Scalable Approach for Decentralized Machine Learning
by: Forootani, Ali, et al.
Published: (2024) -
Synthetic Time Series Forecasting with Transformer Architectures: Extensive Simulation Benchmarks
by: Forootani, Ali, et al.
Published: (2025) -
Learnable Koopman-Enhanced Transformer-Based Time Series Forecasting with Spectral Control
by: Forootani, Ali, et al.
Published: (2026) -
Asynchronous Federated Learning with non-convex client objective functions and heterogeneous dataset
by: Forootani, Ali, et al.
Published: (2025) -
Conformal Off-Policy Evaluation in Markov Decision Processes
by: Foffano, Daniele, et al.
Published: (2023)