Territory Paint Wars: Diagnosing and Mitigating Failure Modes in Competitive Multi-Agent PPO
Fuente:
arXiv
Saved in:
| Main Author: | Singh, Diyansha |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Your Data, My Model: Learning Who Really Helps in Federated Learning
by: Abdurakhmanova, Shamsiiat, et al.
Published: (2024)
by: Abdurakhmanova, Shamsiiat, et al.
Published: (2024)
On the Fundamental Limitations of Decentralized Learnable Reward Shaping in Cooperative Multi-Agent Reinforcement Learning
by: Akella, Aditya
Published: (2025)
by: Akella, Aditya
Published: (2025)
Adaptive Latent-Space Constraints in Personalized Federated Learning
by: Ayromlou, Sana, et al.
Published: (2025)
by: Ayromlou, Sana, et al.
Published: (2025)
MACS: Multi-Agent Reinforcement Learning for Optimization of Crystal Structures
by: Zamaraeva, Elena, et al.
Published: (2025)
by: Zamaraeva, Elena, et al.
Published: (2025)
Decentralized Time Series Classification with ROCKET Features
by: Casella, Bruno, et al.
Published: (2025)
by: Casella, Bruno, et al.
Published: (2025)
Client-Conditional Federated Learning via Local Training Data Statistics
by: Brännvall, Rickard
Published: (2026)
by: Brännvall, Rickard
Published: (2026)
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
by: Pepe, Alberto, et al.
Published: (2026)
by: Pepe, Alberto, et al.
Published: (2026)
DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas
by: Wang, Zhen, et al.
Published: (2025)
by: Wang, Zhen, et al.
Published: (2025)
Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork
by: Jing, Yuheng, et al.
Published: (2026)
by: Jing, Yuheng, et al.
Published: (2026)
Dynamical Priors as a Training Objective in Reinforcement Learning
by: Subaharan, Sukesh
Published: (2026)
by: Subaharan, Sukesh
Published: (2026)
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement
by: Lin, Junjie, et al.
Published: (2024)
by: Lin, Junjie, et al.
Published: (2024)
AI Agents: Evolution, Architecture, and Real-World Applications
by: Krishnan, Naveen
Published: (2025)
by: Krishnan, Naveen
Published: (2025)
Adaptive Minds: Empowering Agents with LoRA-as-Tools
by: Shekar, Pavan C, et al.
Published: (2025)
by: Shekar, Pavan C, et al.
Published: (2025)
Evaluation of Differential Privacy Mechanisms on Federated Learning
by: Varsani, Tejash
Published: (2025)
by: Varsani, Tejash
Published: (2025)
Generative Evolutionary Meta-Solver (GEMS): Scalable Surrogate-Free Multi-Agent Reinforcement Learning
by: Sharma, Alakh, et al.
Published: (2025)
by: Sharma, Alakh, et al.
Published: (2025)
Dynamic Dual-Granularity Skill Bank for Agentic RL
by: Tu, Songjun, et al.
Published: (2026)
by: Tu, Songjun, et al.
Published: (2026)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
by: Jia, Xiao
Published: (2026)
by: Jia, Xiao
Published: (2026)
WorkflowGen:an adaptive workflow generation mechanism driven by trajectory experience
by: Wei, Ruocan, et al.
Published: (2026)
by: Wei, Ruocan, et al.
Published: (2026)
Partitioned Neural Network Training via Synthetic Intermediate Labels
by: Karadağ, Cevat Volkan, et al.
Published: (2024)
by: Karadağ, Cevat Volkan, et al.
Published: (2024)
When Does Global Attention Help? A Unified Empirical Study on Atomistic Graph Learning
by: Chowdhury, Arindam, et al.
Published: (2025)
by: Chowdhury, Arindam, et al.
Published: (2025)
SPARK: Igniting Communication-Efficient Decentralized Learning via Stage-wise Projected NTK and Accelerated Regularization
by: Xia, Li
Published: (2025)
by: Xia, Li
Published: (2025)
Federated Learning and Class Imbalances
by: Zhu, Siqi, et al.
Published: (2026)
by: Zhu, Siqi, et al.
Published: (2026)
FedRot-LoRA: Mitigating Rotational Misalignment in Federated LoRA
by: Zhang, Haoran, et al.
Published: (2026)
by: Zhang, Haoran, et al.
Published: (2026)
Communicating Plans, Not Percepts: Scalable Multi-Agent Coordination with Embodied World Models
by: Hill, Brennen A., et al.
Published: (2025)
by: Hill, Brennen A., et al.
Published: (2025)
How to Correctly do Semantic Backpropagation on Language-based Agentic Systems
by: Wang, Wenyi, et al.
Published: (2024)
by: Wang, Wenyi, et al.
Published: (2024)
A V2X-based Privacy Preserving Federated Measuring and Learning System
by: Alekszejenkó, Levente, et al.
Published: (2024)
by: Alekszejenkó, Levente, et al.
Published: (2024)
Multi-Agent Decision-Focused Learning via Value-Aware Sequential Communication
by: Amoh, Benjamin, et al.
Published: (2026)
by: Amoh, Benjamin, et al.
Published: (2026)
Spatial-Temporal Learning-Based Distributed Routing for Dynamic LEO Satellite Networks
by: Chou, Po-Heng, et al.
Published: (2026)
by: Chou, Po-Heng, et al.
Published: (2026)
The Efficiency Attenuation Phenomenon: A Computational Challenge to the Language of Thought Hypothesis
by: Zhang, Di
Published: (2026)
by: Zhang, Di
Published: (2026)
Manipulating Transformer-Based Models: Controllability, Steerability, and Robust Interventions
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
The Six Sigma Agent: Achieving Enterprise-Grade Reliability in LLM Systems Through Consensus-Driven Decomposed Execution
by: Patel, Khush, et al.
Published: (2026)
by: Patel, Khush, et al.
Published: (2026)
Generating Realistic Safety-Critical Scenarios for Vehicle-Pedestrian Interactions
by: Pu, Qingwen, et al.
Published: (2026)
by: Pu, Qingwen, et al.
Published: (2026)
An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning
by: Saini, Saurabh, et al.
Published: (2026)
by: Saini, Saurabh, et al.
Published: (2026)
MAcPNN: Mutual Assisted Learning on Data Streams with Temporal Dependence
by: Giannini, Federico, et al.
Published: (2026)
by: Giannini, Federico, et al.
Published: (2026)
Semantic-Constrained Federated Aggregation: Convergence Theory and Privacy-Utility Bounds for Knowledge-Enhanced Distributed Learning
by: Arafat, Jahidul
Published: (2025)
by: Arafat, Jahidul
Published: (2025)
RPRA: Predicting an LLM-Judge for Efficient but Performant Inference
by: Ashley, Dylan R., et al.
Published: (2026)
by: Ashley, Dylan R., et al.
Published: (2026)
Correction and Corruption: A Two-Rate View of Error Flow in LLM Protocols
by: Reitich, Fernando
Published: (2026)
by: Reitich, Fernando
Published: (2026)
Distinguished In Uniform: Self Attention Vs. Virtual Nodes
by: Rosenbluth, Eran, et al.
Published: (2024)
by: Rosenbluth, Eran, et al.
Published: (2024)
Coopetition-Gym v1: A Formally Grounded Platform for Mixed-Motive Multi-Agent Reinforcement Learning under Strategic Coopetition
by: Pant, Vik, et al.
Published: (2026)
by: Pant, Vik, et al.
Published: (2026)
The Coordinate System Problem in Persistent Structural Memory for Neural Architectures
by: Basu, Abhinaba
Published: (2026)
by: Basu, Abhinaba
Published: (2026)
Similar Items
-
Your Data, My Model: Learning Who Really Helps in Federated Learning
by: Abdurakhmanova, Shamsiiat, et al.
Published: (2024) -
On the Fundamental Limitations of Decentralized Learnable Reward Shaping in Cooperative Multi-Agent Reinforcement Learning
by: Akella, Aditya
Published: (2025) -
Adaptive Latent-Space Constraints in Personalized Federated Learning
by: Ayromlou, Sana, et al.
Published: (2025) -
MACS: Multi-Agent Reinforcement Learning for Optimization of Crystal Structures
by: Zamaraeva, Elena, et al.
Published: (2025) -
Decentralized Time Series Classification with ROCKET Features
by: Casella, Bruno, et al.
Published: (2025)