Ranking Joint Policies in Dynamic Games using Evolutionary Dynamics
Fuente:
arXiv
Guardado en:
| Autores principales: | Koliou, Natalia, Vouros, George |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Enhancing Cooperative Multi-Agent Reinforcement Learning with State Modelling and Adversarial Exploration
por: Kontogiannis, Andreas, et al.
Publicado: (2025)
por: Kontogiannis, Andreas, et al.
Publicado: (2025)
Learning safe, constrained policies via imitation learning: Connection to Probabilistic Inference and a Naive Algorithm
por: Papadopoulos, George, et al.
Publicado: (2025)
por: Papadopoulos, George, et al.
Publicado: (2025)
Sample-Efficient Policy Space Response Oracles with Joint Experience Best Response
por: Bighashdel, Ariyan, et al.
Publicado: (2026)
por: Bighashdel, Ariyan, et al.
Publicado: (2026)
Trust as Monitoring: Evolutionary Dynamics of User Trust and AI Developer Behaviour
por: Bashir, Adeela, et al.
Publicado: (2026)
por: Bashir, Adeela, et al.
Publicado: (2026)
Towards Learning Scalable Agile Dynamic Motion Planning for Robosoccer Teams with Policy Optimization
por: Ho, Brandon, et al.
Publicado: (2025)
por: Ho, Brandon, et al.
Publicado: (2025)
Dynamic Speculative Agent Planning
por: Guan, Yilin, et al.
Publicado: (2025)
por: Guan, Yilin, et al.
Publicado: (2025)
Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning
por: Zhang, Beining, et al.
Publicado: (2025)
por: Zhang, Beining, et al.
Publicado: (2025)
Solving Continuous Mean Field Games: Deep Reinforcement Learning for Non-Stationary Dynamics
por: Magnino, Lorenzo, et al.
Publicado: (2025)
por: Magnino, Lorenzo, et al.
Publicado: (2025)
Dynamic Sight Range Selection in Multi-Agent Reinforcement Learning
por: Liao, Wei-Chen, et al.
Publicado: (2025)
por: Liao, Wei-Chen, et al.
Publicado: (2025)
Bounded Coupled AI Learning Dynamics in Tri-Hierarchical Drone Swarms
por: Bychkov, Oleksii
Publicado: (2026)
por: Bychkov, Oleksii
Publicado: (2026)
Multi-agent Reinforcement Learning for Dynamic Dispatching in Material Handling Systems
por: Lee, Xian Yeow, et al.
Publicado: (2024)
por: Lee, Xian Yeow, et al.
Publicado: (2024)
LERO: LLM-driven Evolutionary framework with Hybrid Rewards and Enhanced Observation for Multi-Agent Reinforcement Learning
por: Wei, Yuan, et al.
Publicado: (2025)
por: Wei, Yuan, et al.
Publicado: (2025)
Dynamic Pricing in High-Speed Railways Using Multi-Agent Reinforcement Learning
por: Villarrubia-Martin, Enrique Adrian, et al.
Publicado: (2025)
por: Villarrubia-Martin, Enrique Adrian, et al.
Publicado: (2025)
Belief Engine: Configurable and Inspectable Stance Dynamics in Multi-Agent LLM Deliberation
por: Yang, Joshua C., et al.
Publicado: (2026)
por: Yang, Joshua C., et al.
Publicado: (2026)
DynaSwarm: Dynamically Graph Structure Selection for LLM-based Multi-agent System
por: Leong, Hui Yi, et al.
Publicado: (2025)
por: Leong, Hui Yi, et al.
Publicado: (2025)
Quantifying Skill and Chance: A Unified Framework for the Geometry of Games
por: Silver, David H.
Publicado: (2025)
por: Silver, David H.
Publicado: (2025)
Learning Strategy Representation for Imitation Learning in Multi-Agent Games
por: Lei, Shiqi, et al.
Publicado: (2024)
por: Lei, Shiqi, et al.
Publicado: (2024)
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
por: Xu, Zelai, et al.
Publicado: (2023)
por: Xu, Zelai, et al.
Publicado: (2023)
Using Deep Q-Learning to Dynamically Toggle between Push/Pull Actions in Computational Trust Mechanisms
por: Lygizou, Zoi, et al.
Publicado: (2024)
por: Lygizou, Zoi, et al.
Publicado: (2024)
Multi-Agent Decision Transformers for Dynamic Dispatching in Material Handling Systems Leveraging Enterprise Big Data
por: Lee, Xian Yeow, et al.
Publicado: (2024)
por: Lee, Xian Yeow, et al.
Publicado: (2024)
A Decision-Language Model (DLM) for Dynamic Restless Multi-Armed Bandit Tasks in Public Health
por: Behari, Nikhil, et al.
Publicado: (2024)
por: Behari, Nikhil, et al.
Publicado: (2024)
A Benchmark for Multi-Party Negotiation Games from Real Negotiation Data
por: Benac, Leo, et al.
Publicado: (2026)
por: Benac, Leo, et al.
Publicado: (2026)
Quantum Frog: Emergent Cooperation and Difficulty Scaling in a Quantized-Time Cooperative Game
por: Mankarious, Saad
Publicado: (2026)
por: Mankarious, Saad
Publicado: (2026)
Scalable Submodular Policy Optimization via Pruned Submodularity Graph
por: Anand, Aditi, et al.
Publicado: (2025)
por: Anand, Aditi, et al.
Publicado: (2025)
Toward Autonomous Engineering Design: A Knowledge-Guided Multi-Agent Framework
por: Kumar, Varun, et al.
Publicado: (2025)
por: Kumar, Varun, et al.
Publicado: (2025)
Distributed Autonomous Swarm Formation for Dynamic Network Bridging
por: Galliera, Raffaele, et al.
Publicado: (2024)
por: Galliera, Raffaele, et al.
Publicado: (2024)
Network-Constrained Policy Optimization for Adaptive Multi-agent Vehicle Routing
por: Arasteh, Fazel, et al.
Publicado: (2025)
por: Arasteh, Fazel, et al.
Publicado: (2025)
Centralized Permutation Equivariant Policy for Cooperative Multi-Agent Reinforcement Learning
por: Xu, Zhuofan, et al.
Publicado: (2025)
por: Xu, Zhuofan, et al.
Publicado: (2025)
Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning
por: Yang, Shan, et al.
Publicado: (2026)
por: Yang, Shan, et al.
Publicado: (2026)
Deep Reinforcement Learning Agents for Strategic Production Policies in Microeconomic Market Simulations
por: Garrido-Merchán, Eduardo C., et al.
Publicado: (2024)
por: Garrido-Merchán, Eduardo C., et al.
Publicado: (2024)
ToMCAT: Theory-of-Mind for Cooperative Agents in Teams via Multiagent Diffusion Policies
por: Sequeira, Pedro, et al.
Publicado: (2025)
por: Sequeira, Pedro, et al.
Publicado: (2025)
Peer Learning: Learning Complex Policies in Groups from Scratch via Action Recommendations
por: Derstroff, Cedric, et al.
Publicado: (2023)
por: Derstroff, Cedric, et al.
Publicado: (2023)
Fairness for Workers Who Pull the Arms: An Index Based Policy for Allocation of Restless Bandit Tasks
por: Biswas, Arpita, et al.
Publicado: (2023)
por: Biswas, Arpita, et al.
Publicado: (2023)
DCcluster-Opt: Benchmarking Dynamic Multi-Objective Optimization for Geo-Distributed Data Center Workloads
por: Guillen-Perez, Antonio, et al.
Publicado: (2025)
por: Guillen-Perez, Antonio, et al.
Publicado: (2025)
Prompting Policies for Multi-step Reasoning and Tool-Use in Black-box LLMs with Iterative Distillation of Experience
por: Sayana, Krishna, et al.
Publicado: (2026)
por: Sayana, Krishna, et al.
Publicado: (2026)
ATHENA: Agentic Team for Hierarchical Evolutionary Numerical Algorithms
por: Toscano, Juan Diego, et al.
Publicado: (2025)
por: Toscano, Juan Diego, et al.
Publicado: (2025)
Generalized Information Gathering Under Dynamics Uncertainty
por: Palafox, Fernando, et al.
Publicado: (2026)
por: Palafox, Fernando, et al.
Publicado: (2026)
Advancing Agentic Systems: Dynamic Task Decomposition, Tool Integration and Evaluation using Novel Metrics and Dataset
por: Gabriel, Adrian Garret, et al.
Publicado: (2024)
por: Gabriel, Adrian Garret, et al.
Publicado: (2024)
Hierarchical LLM-Driven Control for HAPS-Assisted UAV Networks: Joint Optimization of Flight and Connectivity
por: Yan, Zijiang, et al.
Publicado: (2026)
por: Yan, Zijiang, et al.
Publicado: (2026)
Distributed Multi-Agent Reinforcement Learning Based on Graph-Induced Local Value Functions
por: Jing, Gangshan, et al.
Publicado: (2022)
por: Jing, Gangshan, et al.
Publicado: (2022)
Ejemplares similares
-
Enhancing Cooperative Multi-Agent Reinforcement Learning with State Modelling and Adversarial Exploration
por: Kontogiannis, Andreas, et al.
Publicado: (2025) -
Learning safe, constrained policies via imitation learning: Connection to Probabilistic Inference and a Naive Algorithm
por: Papadopoulos, George, et al.
Publicado: (2025) -
Sample-Efficient Policy Space Response Oracles with Joint Experience Best Response
por: Bighashdel, Ariyan, et al.
Publicado: (2026) -
Trust as Monitoring: Evolutionary Dynamics of User Trust and AI Developer Behaviour
por: Bashir, Adeela, et al.
Publicado: (2026) -
Towards Learning Scalable Agile Dynamic Motion Planning for Robosoccer Teams with Policy Optimization
por: Ho, Brandon, et al.
Publicado: (2025)