Risk-seeking conservative policy iteration with agent-state based policies for Dec-POMDPs with guaranteed convergence
Fuente:
arXiv
Guardado en:
| Autores principales: | Sinha, Amit, Geist, Matthieu, Mahajan, Aditya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Periodic agent-state based Q-learning for POMDPs
por: Sinha, Amit, et al.
Publicado: (2024)
por: Sinha, Amit, et al.
Publicado: (2024)
Agent-state based policies in POMDPs: Beyond belief-state MDPs
por: Sinha, Amit, et al.
Publicado: (2024)
por: Sinha, Amit, et al.
Publicado: (2024)
Convergence of regularized agent-state-based Q-learning in POMDPs
por: Sinha, Amit, et al.
Publicado: (2025)
por: Sinha, Amit, et al.
Publicado: (2025)
High entropy leads to symmetry equivariant policies in Dec-POMDPs
por: Forkel, Johannes, et al.
Publicado: (2025)
por: Forkel, Johannes, et al.
Publicado: (2025)
Solving Collaborative Dec-POMDPs with Deep Reinforcement Learning Heuristics
por: Soffair, Nitsan
Publicado: (2022)
por: Soffair, Nitsan
Publicado: (2022)
Kernel-based learning with guarantees for multi-agent applications
por: Kowalczyk, Krzysztof, et al.
Publicado: (2024)
por: Kowalczyk, Krzysztof, et al.
Publicado: (2024)
Towards Optimal Performance and Action Consistency Guarantees in Dec-POMDPs with Inconsistent Beliefs and Limited Communication
por: Shimron, Moshe Rafaeli, et al.
Publicado: (2025)
por: Shimron, Moshe Rafaeli, et al.
Publicado: (2025)
Distributed scalable coupled policy algorithm for networked multi-agent reinforcement learning
por: Dai, Pengcheng, et al.
Publicado: (2025)
por: Dai, Pengcheng, et al.
Publicado: (2025)
Distributed primal-dual algorithm for constrained multi-agent reinforcement learning under coupled policies
por: Dai, Pengcheng, et al.
Publicado: (2025)
por: Dai, Pengcheng, et al.
Publicado: (2025)
Optimally Solving Simultaneous-Move Dec-POMDPs: The Sequential Central Planning Approach
por: Peralez, Johan, et al.
Publicado: (2024)
por: Peralez, Johan, et al.
Publicado: (2024)
Learning to Deliberate: Meta-policy Collaboration for Agentic LLMs with Multi-agent Reinforcement Learning
por: Yang, Wei, et al.
Publicado: (2025)
por: Yang, Wei, et al.
Publicado: (2025)
Multi-agent Off-policy Actor-Critic Reinforcement Learning for Partially Observable Environments
por: Zhaikhan, Ainur, et al.
Publicado: (2024)
por: Zhaikhan, Ainur, et al.
Publicado: (2024)
Factored Online Planning in Many-Agent POMDPs
por: Galesloot, Maris F. L., et al.
Publicado: (2023)
por: Galesloot, Maris F. L., et al.
Publicado: (2023)
Population-aware Online Mirror Descent for Mean-Field Games with Common Noise by Deep Reinforcement Learning
por: Wu, Zida, et al.
Publicado: (2025)
por: Wu, Zida, et al.
Publicado: (2025)
Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient
por: Yu, Xiaoyang, et al.
Publicado: (2025)
por: Yu, Xiaoyang, et al.
Publicado: (2025)
Quantized distributed Nash equilibrium seeking under DoS attacks
por: Feng, Shuai, et al.
Publicado: (2023)
por: Feng, Shuai, et al.
Publicado: (2023)
Probing Dec-POMDP Reasoning in Cooperative MARL
por: Tessera, Kale-ab, et al.
Publicado: (2026)
por: Tessera, Kale-ab, et al.
Publicado: (2026)
Modeling human reputation-seeking behavior in a spatio-temporally complex public good provision game
por: Hughes, Edward, et al.
Publicado: (2025)
por: Hughes, Edward, et al.
Publicado: (2025)
Consensus seeking in diffusive multidimensional networks with a repeated interaction pattern and time-delays
por: Vu, Hoang Huy, et al.
Publicado: (2024)
por: Vu, Hoang Huy, et al.
Publicado: (2024)
On-policy Actor-Critic Reinforcement Learning for Multi-UAV Exploration
por: Farid, Ali Moltajaei, et al.
Publicado: (2024)
por: Farid, Ali Moltajaei, et al.
Publicado: (2024)
Bench-MFG: A Benchmark Suite for Learning in Stationary Mean Field Games
por: Magnino, Lorenzo, et al.
Publicado: (2026)
por: Magnino, Lorenzo, et al.
Publicado: (2026)
Greedy-based Value Representation for Optimal Coordination in Multi-agent Reinforcement Learning
por: Wan, Lipeng, et al.
Publicado: (2021)
por: Wan, Lipeng, et al.
Publicado: (2021)
Towards a satisfactory conversion of messages among agent-based information systems
por: Berges, Idoia, et al.
Publicado: (2024)
por: Berges, Idoia, et al.
Publicado: (2024)
Replicating the behaviour of electric vehicle drivers using an agent-based reinforcement learning model
por: Feng, Zixin, et al.
Publicado: (2025)
por: Feng, Zixin, et al.
Publicado: (2025)
Middleware-based multi-agent development environment for building and testing distributed intelligent systems
por: Aguayo-Canela, Francisco José, et al.
Publicado: (2024)
por: Aguayo-Canela, Francisco José, et al.
Publicado: (2024)
DHLight: Multi-agent Policy-based Directed Hypergraph Learning for Traffic Signal Control
por: Lei, Zhen, et al.
Publicado: (2024)
por: Lei, Zhen, et al.
Publicado: (2024)
A dynamic state-based model of crowds
por: Amos, Martyn, et al.
Publicado: (2023)
por: Amos, Martyn, et al.
Publicado: (2023)
Enhancing healthcare infrastructure resilience through agent-based simulation methods
por: Carramiñana, David, et al.
Publicado: (2025)
por: Carramiñana, David, et al.
Publicado: (2025)
Collaborative Threat-Aware Autonomy (CTAA)
por: Sharma, Rajnikant, et al.
Publicado: (2026)
por: Sharma, Rajnikant, et al.
Publicado: (2026)
Enriched multi-agent middleware for building rule-based distributed security solutions for IoT environments
por: Aguayo-Canela, Francisco José, et al.
Publicado: (2024)
por: Aguayo-Canela, Francisco José, et al.
Publicado: (2024)
Finite-time convergence to an $ε$-efficient Nash equilibrium in potential games
por: Maddux, Anna, et al.
Publicado: (2024)
por: Maddux, Anna, et al.
Publicado: (2024)
A low-cost Framework for Decentralized Autonomous Intersection Management
por: Katole, Rugved, et al.
Publicado: (2023)
por: Katole, Rugved, et al.
Publicado: (2023)
Multi-agent Reinforcement Traffic Signal Control based on Interpretable Influence Mechanism and Biased ReLU Approximation
por: Luo, Zhiyue, et al.
Publicado: (2024)
por: Luo, Zhiyue, et al.
Publicado: (2024)
Multi-agent based modeling for investigating excess heat utilization from electrolyzer production to district heating network
por: Christensen, Kristoffer, et al.
Publicado: (2024)
por: Christensen, Kristoffer, et al.
Publicado: (2024)
Critical mobility in policy making for epidemic containment
por: López, Jesús A. Moreno, et al.
Publicado: (2024)
por: López, Jesús A. Moreno, et al.
Publicado: (2024)
On the limits of agency in agent-based models
por: Chopra, Ayush, et al.
Publicado: (2024)
por: Chopra, Ayush, et al.
Publicado: (2024)
Nash equilibrium seeking for a class of quadratic-bilinear Wasserstein distributionally robust games
por: Pantazis, Georgios, et al.
Publicado: (2024)
por: Pantazis, Georgios, et al.
Publicado: (2024)
Teaching an Old Dynamics New Tricks: Regularization-free Last-iterate Convergence in Zero-sum Games via BNN Dynamics
por: Zhang, Tuo, et al.
Publicado: (2026)
por: Zhang, Tuo, et al.
Publicado: (2026)
Joint Optimization of Multi-agent Memory System
por: Mao, Wenyu, et al.
Publicado: (2026)
por: Mao, Wenyu, et al.
Publicado: (2026)
Learning safe, constrained policies via imitation learning: Connection to Probabilistic Inference and a Naive Algorithm
por: Papadopoulos, George, et al.
Publicado: (2025)
por: Papadopoulos, George, et al.
Publicado: (2025)
Ejemplares similares
-
Periodic agent-state based Q-learning for POMDPs
por: Sinha, Amit, et al.
Publicado: (2024) -
Agent-state based policies in POMDPs: Beyond belief-state MDPs
por: Sinha, Amit, et al.
Publicado: (2024) -
Convergence of regularized agent-state-based Q-learning in POMDPs
por: Sinha, Amit, et al.
Publicado: (2025) -
High entropy leads to symmetry equivariant policies in Dec-POMDPs
por: Forkel, Johannes, et al.
Publicado: (2025) -
Solving Collaborative Dec-POMDPs with Deep Reinforcement Learning Heuristics
por: Soffair, Nitsan
Publicado: (2022)