Generating Local Shields for Decentralised Partially Observable Markov Decision Processes
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Haoran, Yoshida, Nobuko |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Nash Approximation Gap in Truncated Infinite-horizon Partially Observable Markov Games
por: Sang, Lan, et al.
Publicado: (2026)
por: Sang, Lan, et al.
Publicado: (2026)
Internal State-Based Policy Gradient Methods for Partially Observable Markov Potential Games
por: Yang, Wonseok, et al.
Publicado: (2026)
por: Yang, Wonseok, et al.
Publicado: (2026)
Safe and Efficient CAV Lane Changing using Decentralised Safety Shields
por: Hegde, Bharathkumar, et al.
Publicado: (2025)
por: Hegde, Bharathkumar, et al.
Publicado: (2025)
Policy Optimization in Multi-Agent Settings under Partially Observable Environments
por: Zhaikhan, Ainur, et al.
Publicado: (2025)
por: Zhaikhan, Ainur, et al.
Publicado: (2025)
Dynamic Graph Communication for Decentralised Multi-Agent Reinforcement Learning
por: McClusky, Ben
Publicado: (2024)
por: McClusky, Ben
Publicado: (2024)
Networked Communication for Decentralised Cooperative Agents in Mean-Field Control
por: Benjamin, Patrick, et al.
Publicado: (2025)
por: Benjamin, Patrick, et al.
Publicado: (2025)
MACRO-LLM: LLM-Empowered Multi-Agent Collaborative Reasoning under Spatiotemporal Partial Observability
por: Chen, Handi, et al.
Publicado: (2026)
por: Chen, Handi, et al.
Publicado: (2026)
Evader-Agnostic Team-Based Pursuit Strategies in Partially-Observable Environments
por: Kalanther, Addison, et al.
Publicado: (2025)
por: Kalanther, Addison, et al.
Publicado: (2025)
Decentralised multi-agent coordination for real-time railway traffic management
por: D'Amato, Leo, et al.
Publicado: (2025)
por: D'Amato, Leo, et al.
Publicado: (2025)
PIANIST: Learning Partially Observable World Models with LLMs for Multi-Agent Decision Making
por: Light, Jonathan, et al.
Publicado: (2024)
por: Light, Jonathan, et al.
Publicado: (2024)
LLaMAR: Long-Horizon Planning for Multi-Agent Robots in Partially Observable Environments
por: Nayak, Siddharth, et al.
Publicado: (2024)
por: Nayak, Siddharth, et al.
Publicado: (2024)
Coordinated Autonomous Drones for Human-Centered Fire Evacuation in Partially Observable Urban Environments
por: Mendoza, Maria G., et al.
Publicado: (2025)
por: Mendoza, Maria G., et al.
Publicado: (2025)
Independent Learning of Nash Equilibria in Partially Observable Markov Potential Games with Decoupled Dynamics
por: Jordan, Philip, et al.
Publicado: (2026)
por: Jordan, Philip, et al.
Publicado: (2026)
On Mobile Ad Hoc Networks for Coverage of Partially Observable Worlds
por: Meriaux, Edwin, et al.
Publicado: (2025)
por: Meriaux, Edwin, et al.
Publicado: (2025)
Multi-agent Off-policy Actor-Critic Reinforcement Learning for Partially Observable Environments
por: Zhaikhan, Ainur, et al.
Publicado: (2024)
por: Zhaikhan, Ainur, et al.
Publicado: (2024)
Computing Universal Plans for Partially Observable Multi-Agent Routing Using Answer Set Programming
por: Zhu, Fengming, et al.
Publicado: (2023)
por: Zhu, Fengming, et al.
Publicado: (2023)
Transparency as Delayed Observability in Multi-Agent Systems
por: Dwarakanath, Kshama, et al.
Publicado: (2024)
por: Dwarakanath, Kshama, et al.
Publicado: (2024)
The Partially Observable Off-Switch Game
por: Garber, Andrew, et al.
Publicado: (2024)
por: Garber, Andrew, et al.
Publicado: (2024)
Approximate Linear Programming for Decentralized Policy Iteration in Cooperative Multi-agent Markov Decision Processes
por: Mandal, Lakshmi, et al.
Publicado: (2023)
por: Mandal, Lakshmi, et al.
Publicado: (2023)
Learning Decentralized Partially Observable Mean Field Control for Artificial Collective Behavior
por: Cui, Kai, et al.
Publicado: (2023)
por: Cui, Kai, et al.
Publicado: (2023)
Asynchronous Training of Mixed-Role Human Actors in a Partially-Observable Environment
por: Chang, Kimberlee Chestnut, et al.
Publicado: (2024)
por: Chang, Kimberlee Chestnut, et al.
Publicado: (2024)
Generalized Coordination of Partially Cooperative Urban Traffic
por: Mertens, Max Bastian, et al.
Publicado: (2025)
por: Mertens, Max Bastian, et al.
Publicado: (2025)
RecBayes: Recurrent Bayesian Ad Hoc Teamwork in Large Partially Observable Domains
por: Ribeiro, João G., et al.
Publicado: (2025)
por: Ribeiro, João G., et al.
Publicado: (2025)
Partially Observable Multi-Agent Reinforcement Learning with Information Sharing
por: Liu, Xiangyu, et al.
Publicado: (2023)
por: Liu, Xiangyu, et al.
Publicado: (2023)
From General Relation Patterns to Task-Specific Decision-Making in Continual Multi-Agent Coordination
por: Yao, Chang, et al.
Publicado: (2025)
por: Yao, Chang, et al.
Publicado: (2025)
Evolution of Social Norms in LLM Agents using Natural Language
por: Horiguchi, Ilya, et al.
Publicado: (2024)
por: Horiguchi, Ilya, et al.
Publicado: (2024)
Observation Interference in Partially Observable Assistance Games
por: Emmons, Scott, et al.
Publicado: (2024)
por: Emmons, Scott, et al.
Publicado: (2024)
Pragmatic Communication for Remote Control of Finite-State Markov Processes
por: Talli, Pietro, et al.
Publicado: (2024)
por: Talli, Pietro, et al.
Publicado: (2024)
Effects of Property Recovery Incentives and Social Interaction on Self-Evacuation Decisions in Natural Disasters: An Agent-Based Modelling Approach
por: Krisnanda, Made, et al.
Publicado: (2026)
por: Krisnanda, Made, et al.
Publicado: (2026)
Online Competitive Information Gathering for Partially Observable Trajectory Games
por: Krusniak, Mel, et al.
Publicado: (2025)
por: Krusniak, Mel, et al.
Publicado: (2025)
Networked Communication for Decentralised Agents in Mean-Field Games
por: Benjamin, Patrick, et al.
Publicado: (2023)
por: Benjamin, Patrick, et al.
Publicado: (2023)
Autonomous Decision Making for Air Taxi Networks
por: Vesel, Alex
Publicado: (2024)
por: Vesel, Alex
Publicado: (2024)
CogSearch: A Cognitive-Aligned Multi-Agent Framework for Proactive Decision Support in E-Commerce Search
por: Zhai, Zhouwei, et al.
Publicado: (2026)
por: Zhai, Zhouwei, et al.
Publicado: (2026)
Decision-Level Fusion for Robust Wearable Affect Recognition
por: Singh, Lokesh, et al.
Publicado: (2026)
por: Singh, Lokesh, et al.
Publicado: (2026)
rAIson: Developing Reliable Decision-Making Agents
por: Moraitis, Pavlos, et al.
Publicado: (2026)
por: Moraitis, Pavlos, et al.
Publicado: (2026)
Partial Resilient Leader-Follower Consensus in Time-Varying Graphs
por: Lee, Haejoon, et al.
Publicado: (2025)
por: Lee, Haejoon, et al.
Publicado: (2025)
Hierarchical Decision-Making in Population Games
por: Chen, Yu-Wen, et al.
Publicado: (2025)
por: Chen, Yu-Wen, et al.
Publicado: (2025)
Beyond Black-Box Benchmarking: Observability, Analytics, and Optimization of Agentic Systems
por: Moshkovich, Dany, et al.
Publicado: (2025)
por: Moshkovich, Dany, et al.
Publicado: (2025)
Grounded Answers for Multi-agent Decision-making Problem through Generative World Model
por: Liu, Zeyang, et al.
Publicado: (2024)
por: Liu, Zeyang, et al.
Publicado: (2024)
Evidence-Decision-Feedback: Theory-Driven Adaptive Scaffolding for LLM Agents
por: Cohn, Clayton, et al.
Publicado: (2026)
por: Cohn, Clayton, et al.
Publicado: (2026)
Ejemplares similares
-
Nash Approximation Gap in Truncated Infinite-horizon Partially Observable Markov Games
por: Sang, Lan, et al.
Publicado: (2026) -
Internal State-Based Policy Gradient Methods for Partially Observable Markov Potential Games
por: Yang, Wonseok, et al.
Publicado: (2026) -
Safe and Efficient CAV Lane Changing using Decentralised Safety Shields
por: Hegde, Bharathkumar, et al.
Publicado: (2025) -
Policy Optimization in Multi-Agent Settings under Partially Observable Environments
por: Zhaikhan, Ainur, et al.
Publicado: (2025) -
Dynamic Graph Communication for Decentralised Multi-Agent Reinforcement Learning
por: McClusky, Ben
Publicado: (2024)