Continuous-Time Distributed Dynamic Programming for Networked Multi-Agent Markov Decision Processes
Fuente:
arXiv
Guardado en:
| Autores principales: | Lee, Donghwan, Lim, Han-Dong, Kim, Do Wan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Interval Markov Decision Processes with Continuous Action-Spaces
por: Delimpaltadakis, Giannis, et al.
Publicado: (2022)
por: Delimpaltadakis, Giannis, et al.
Publicado: (2022)
Harnessing Membership Function Dynamics for Stability Analysis of T-S Fuzzy Systems
por: Lee, Donghwan, et al.
Publicado: (2024)
por: Lee, Donghwan, et al.
Publicado: (2024)
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
por: Lee, Donghwan
Publicado: (2026)
por: Lee, Donghwan
Publicado: (2026)
Intermittently Observable Markov Decision Processes
por: Chen, Gongpu, et al.
Publicado: (2023)
por: Chen, Gongpu, et al.
Publicado: (2023)
Causal Temporal Reasoning for Markov Decision Processes
por: Kazemi, Milad, et al.
Publicado: (2022)
por: Kazemi, Milad, et al.
Publicado: (2022)
OCMDP: Observation-Constrained Markov Decision Process
por: Wang, Taiyi, et al.
Publicado: (2024)
por: Wang, Taiyi, et al.
Publicado: (2024)
Finite-Time Accuracy of Temporal-Difference Learning Under Schur-Stable Recursions
por: Lee, Donghwan, et al.
Publicado: (2022)
por: Lee, Donghwan, et al.
Publicado: (2022)
Lyapunov-Certified Direct Switching Theory for Q-Learning
por: Lee, Donghwan
Publicado: (2026)
por: Lee, Donghwan
Publicado: (2026)
Optimal Control of Logically Constrained Partially Observable and Multi-Agent Markov Decision Processes
por: Kalagarla, Krishna C., et al.
Publicado: (2023)
por: Kalagarla, Krishna C., et al.
Publicado: (2023)
Optimized Task Assignment and Predictive Maintenance for Industrial Machines using Markov Decision Process
por: Nasir, Ali, et al.
Publicado: (2024)
por: Nasir, Ali, et al.
Publicado: (2024)
Conformal Off-Policy Evaluation in Markov Decision Processes
por: Foffano, Daniele, et al.
Publicado: (2023)
por: Foffano, Daniele, et al.
Publicado: (2023)
Learning Algorithms for Verification of Markov Decision Processes
por: Brázdil, Tomáš, et al.
Publicado: (2024)
por: Brázdil, Tomáš, et al.
Publicado: (2024)
A finite time analysis of distributed Q-learning
por: Lim, Han-Dong, et al.
Publicado: (2024)
por: Lim, Han-Dong, et al.
Publicado: (2024)
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
por: Angelotti, Giorgio, et al.
Publicado: (2021)
por: Angelotti, Giorgio, et al.
Publicado: (2021)
On Dynamic Programming Decompositions of Static Risk Measures in Markov Decision Processes
por: Hau, Jia Lin, et al.
Publicado: (2023)
por: Hau, Jia Lin, et al.
Publicado: (2023)
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
por: Murthy, Yashaswini, et al.
Publicado: (2023)
por: Murthy, Yashaswini, et al.
Publicado: (2023)
Towards Hierarchical Multi-Agent Decision-Making for Uncertainty-Aware EV Charging
por: Ting, Lo Pang-Yun, et al.
Publicado: (2024)
por: Ting, Lo Pang-Yun, et al.
Publicado: (2024)
Performance-Aware Self-Configurable Multi-Agent Networks: A Distributed Submodular Approach for Simultaneous Coordination and Network Design
por: Xu, Zirui, et al.
Publicado: (2024)
por: Xu, Zirui, et al.
Publicado: (2024)
Optimizing Return Distributions with Distributional Dynamic Programming
por: Pires, Bernardo Ávila, et al.
Publicado: (2025)
por: Pires, Bernardo Ávila, et al.
Publicado: (2025)
Nonuniform-Grid Markov Chain Approximation of Continuous Processes with Time-Linear Moments
por: Kim, Do Hyun, et al.
Publicado: (2025)
por: Kim, Do Hyun, et al.
Publicado: (2025)
Verification of Neural Network Control Systems in Continuous Time
por: ArjomandBigdeli, Ali, et al.
Publicado: (2024)
por: ArjomandBigdeli, Ali, et al.
Publicado: (2024)
A Distributed Hierarchical Spatio-Temporal Edge-Enhanced Graph Neural Network for City-Scale Dynamic Logistics Routing
por: Han, Zihan, et al.
Publicado: (2025)
por: Han, Zihan, et al.
Publicado: (2025)
Resource-Aware Distributed Submodular Maximization: A Paradigm for Multi-Robot Decision-Making
por: Xu, Zirui, et al.
Publicado: (2022)
por: Xu, Zirui, et al.
Publicado: (2022)
Distributionally Robust Free Energy Principle for Decision-Making
por: Shafiei, Allahkaram, et al.
Publicado: (2025)
por: Shafiei, Allahkaram, et al.
Publicado: (2025)
Challenges in Applying Variational Quantum Algorithms to Dynamic Satellite Network Routing
por: Do, Phuc Hao, et al.
Publicado: (2025)
por: Do, Phuc Hao, et al.
Publicado: (2025)
Dispatch-Aware Deep Neural Network for Optimal Transmission Switching: Toward Real-Time and Feasibility Guaranteed Operation
por: Kim, Minsoo, et al.
Publicado: (2025)
por: Kim, Minsoo, et al.
Publicado: (2025)
Customized User Plane Processing via Code Generating AI Agents for Next Generation Mobile Networks
por: Ma, Xiaowen, et al.
Publicado: (2026)
por: Ma, Xiaowen, et al.
Publicado: (2026)
Switching-Geometry Analysis of Deflated Q-Value Iteration
por: Lee, Donghwan
Publicado: (2026)
por: Lee, Donghwan
Publicado: (2026)
Improved Monte Carlo Planning via Causal Disentanglement for Structurally-Decomposed Markov Decision Processes
por: Liu, Larkin, et al.
Publicado: (2024)
por: Liu, Larkin, et al.
Publicado: (2024)
Neural Continuous-Time Supermartingale Certificates
por: Neustroev, Grigory, et al.
Publicado: (2024)
por: Neustroev, Grigory, et al.
Publicado: (2024)
Large Language Models for Explainable Decisions in Dynamic Digital Twins
por: Zhang, Nan, et al.
Publicado: (2024)
por: Zhang, Nan, et al.
Publicado: (2024)
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
por: Azeem, Muqsit, et al.
Publicado: (2024)
por: Azeem, Muqsit, et al.
Publicado: (2024)
Exploring a Physics-Informed Decision Transformer for Distribution System Restoration: Methodology and Performance Analysis
por: Zhao, Hong, et al.
Publicado: (2024)
por: Zhao, Hong, et al.
Publicado: (2024)
HONEST-CAV: Hierarchical Optimization of Network Signals and Trajectories for Connected and Automated Vehicles with Multi-Agent Reinforcement Learning
por: Zhang, Ziyan, et al.
Publicado: (2026)
por: Zhang, Ziyan, et al.
Publicado: (2026)
GenSafe: A Generalizable Safety Enhancer for Safe Reinforcement Learning Algorithms Based on Reduced Order Markov Decision Process Model
por: Zhou, Zhehua, et al.
Publicado: (2024)
por: Zhou, Zhehua, et al.
Publicado: (2024)
Trusted Routing for Blockchain-Empowered UAV Networks via Multi-Agent Deep Reinforcement Learning
por: Jia, Ziye, et al.
Publicado: (2025)
por: Jia, Ziye, et al.
Publicado: (2025)
Safe Decentralized Operation of EV Virtual Power Plant with Limited Network Visibility via Multi-Agent Reinforcement Learning
por: Huang, Chenghao, et al.
Publicado: (2026)
por: Huang, Chenghao, et al.
Publicado: (2026)
Perimeter Control with Heterogeneous Metering Rates for Cordon Signals: A Physics-Regularized Multi-Agent Reinforcement Learning Approach
por: Yu, Jiajie, et al.
Publicado: (2023)
por: Yu, Jiajie, et al.
Publicado: (2023)
Finite-Time Analysis of Temporal Difference Learning with Experience Replay
por: Lim, Han-Dong, et al.
Publicado: (2023)
por: Lim, Han-Dong, et al.
Publicado: (2023)
Fine-Tuning and Prompt Engineering of LLMs, for the Creation of Multi-Agent AI for Addressing Sustainable Protein Production Challenges
por: Kalian, Alexander D., et al.
Publicado: (2025)
por: Kalian, Alexander D., et al.
Publicado: (2025)
Ejemplares similares
-
Interval Markov Decision Processes with Continuous Action-Spaces
por: Delimpaltadakis, Giannis, et al.
Publicado: (2022) -
Harnessing Membership Function Dynamics for Stability Analysis of T-S Fuzzy Systems
por: Lee, Donghwan, et al.
Publicado: (2024) -
Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration
por: Lee, Donghwan
Publicado: (2026) -
Intermittently Observable Markov Decision Processes
por: Chen, Gongpu, et al.
Publicado: (2023) -
Causal Temporal Reasoning for Markov Decision Processes
por: Kazemi, Milad, et al.
Publicado: (2022)