Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lauffer, Niklas, Shah, Ameesh, Carroll, Micah, Seshia, Sanjit A., Russell, Stuart, Dennis, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
von: Tang, Wenjie, et al.
Veröffentlicht: (2026)
von: Tang, Wenjie, et al.
Veröffentlicht: (2026)
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
von: Castellini, Jacopo, et al.
Veröffentlicht: (2019)
von: Castellini, Jacopo, et al.
Veröffentlicht: (2019)
ME-IGM: Individual-Global-Max in Maximum Entropy Multi-Agent Reinforcement Learning
von: Chen, Wen-Tse, et al.
Veröffentlicht: (2024)
von: Chen, Wen-Tse, et al.
Veröffentlicht: (2024)
Difference Rewards Policy Gradients
von: Castellini, Jacopo, et al.
Veröffentlicht: (2020)
von: Castellini, Jacopo, et al.
Veröffentlicht: (2020)
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
von: Pasupuleti, Vinil, et al.
Veröffentlicht: (2026)
von: Pasupuleti, Vinil, et al.
Veröffentlicht: (2026)
One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents
von: Hong, Yoosung
Veröffentlicht: (2026)
von: Hong, Yoosung
Veröffentlicht: (2026)
ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation
von: Mittal, Tarun
Veröffentlicht: (2026)
von: Mittal, Tarun
Veröffentlicht: (2026)
PARNESS: A Paper Harness for End-to-End Automated Scientific Research with Dynamic Workflows, Full-Text Indexing, and Cross-Run Knowledge Accumulation
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
von: Wang, Yuchen, et al.
Veröffentlicht: (2026)
When Agents Disagree: The Selection Bottleneck in Multi-Agent LLM Pipelines
von: Maryanskyy, Artem
Veröffentlicht: (2026)
von: Maryanskyy, Artem
Veröffentlicht: (2026)
Extending NGU to Multi-Agent RL: A Preliminary Study
von: Hernandez, Juan, et al.
Veröffentlicht: (2025)
von: Hernandez, Juan, et al.
Veröffentlicht: (2025)
Advancing Multimodal Agent Reasoning with Long-Term Neuro-Symbolic Memory
von: Jiang, Rongjie, et al.
Veröffentlicht: (2026)
von: Jiang, Rongjie, et al.
Veröffentlicht: (2026)
Instruction-Level Weight Shaping: A Framework for Self-Improving AI Agents
von: Costa, Rimom
Veröffentlicht: (2025)
von: Costa, Rimom
Veröffentlicht: (2025)
Learning To Help: Training Models to Assist Legacy Devices
von: Wu, Yu, et al.
Veröffentlicht: (2024)
von: Wu, Yu, et al.
Veröffentlicht: (2024)
StatePlane: A Cognitive State Plane for Long-Horizon AI Systems Under Bounded Context
von: Annapureddy, Sasank, et al.
Veröffentlicht: (2026)
von: Annapureddy, Sasank, et al.
Veröffentlicht: (2026)
QTypeMix: Enhancing Multi-Agent Cooperative Strategies through Heterogeneous and Homogeneous Value Decomposition
von: Fu, Songchen, et al.
Veröffentlicht: (2024)
von: Fu, Songchen, et al.
Veröffentlicht: (2024)
FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing
von: Zhang, Mingda, et al.
Veröffentlicht: (2026)
von: Zhang, Mingda, et al.
Veröffentlicht: (2026)
Representational Collapse in Multi-Agent LLM Committees: Measurement and Diversity-Aware Consensus
von: Patel, Dipkumar
Veröffentlicht: (2026)
von: Patel, Dipkumar
Veröffentlicht: (2026)
Bimanual Robot Manipulation via Multi-Agent In-Context Learning
von: Palma, Alessio, et al.
Veröffentlicht: (2026)
von: Palma, Alessio, et al.
Veröffentlicht: (2026)
N-Agent Ad Hoc Teamwork
von: Wang, Caroline, et al.
Veröffentlicht: (2024)
von: Wang, Caroline, et al.
Veröffentlicht: (2024)
Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Bias
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
von: Wu, Shuai, et al.
Veröffentlicht: (2026)
Enhancing Heterogeneous Multi-Agent Cooperation in Decentralized MARL via GNN-driven Intrinsic Rewards
von: Monon, Jahir Sadik, et al.
Veröffentlicht: (2024)
von: Monon, Jahir Sadik, et al.
Veröffentlicht: (2024)
Latent Cache Flow: Model-to-Model Communication Without Text
von: Rossi, Maximillian, et al.
Veröffentlicht: (2026)
von: Rossi, Maximillian, et al.
Veröffentlicht: (2026)
A Framework for Assessing AI Agent Decisions and Outcomes in AutoML Pipelines
von: Du, Gaoyuan, et al.
Veröffentlicht: (2026)
von: Du, Gaoyuan, et al.
Veröffentlicht: (2026)
Umwelt Engineering: Designing the Cognitive Worlds of Linguistic Agents
von: Jehu-Appiah, Rodney
Veröffentlicht: (2026)
von: Jehu-Appiah, Rodney
Veröffentlicht: (2026)
PillagerBench: Benchmarking LLM-Based Agents in Competitive Minecraft Team Environments
von: Schipper, Olivier, et al.
Veröffentlicht: (2025)
von: Schipper, Olivier, et al.
Veröffentlicht: (2025)
Dynamic Attentional Context Scoping: Agent-Triggered Focus Sessions for Isolated Per-Agent Steering in Multi-Agent LLM Orchestration
von: Patel, Nickson
Veröffentlicht: (2026)
von: Patel, Nickson
Veröffentlicht: (2026)
Collaboration Promotes Group Resilience in Multi-Agent RL
von: Shraga, Ilai, et al.
Veröffentlicht: (2021)
von: Shraga, Ilai, et al.
Veröffentlicht: (2021)
Decentralized Aerial Manipulation of a Cable-Suspended Load using Multi-Agent Reinforcement Learning
von: Zeng, Jack, et al.
Veröffentlicht: (2025)
von: Zeng, Jack, et al.
Veröffentlicht: (2025)
A Principle of Targeted Intervention for Multi-Agent Reinforcement Learning
von: Liu, Anjie, et al.
Veröffentlicht: (2025)
von: Liu, Anjie, et al.
Veröffentlicht: (2025)
MACS: Multi-Agent Reinforcement Learning for Optimization of Crystal Structures
von: Zamaraeva, Elena, et al.
Veröffentlicht: (2025)
von: Zamaraeva, Elena, et al.
Veröffentlicht: (2025)
Policy Search, Retrieval, and Composition via Task Similarity in Collaborative Agentic Systems
von: Nath, Saptarshi, et al.
Veröffentlicht: (2025)
von: Nath, Saptarshi, et al.
Veröffentlicht: (2025)
Knowledge Equivalence in Digital Twins of Intelligent Systems
von: Zhang, Nan, et al.
Veröffentlicht: (2022)
von: Zhang, Nan, et al.
Veröffentlicht: (2022)
Risk-Sensitive Multi-Agent Reinforcement Learning in Network Aggregative Markov Games
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2024)
von: Ghaemi, Hafez, et al.
Veröffentlicht: (2024)
Differentiable Model Predictive Safety for Heterogeneous Mobility at Urban Intersections
von: Song, Wenzhe, et al.
Veröffentlicht: (2026)
von: Song, Wenzhe, et al.
Veröffentlicht: (2026)
AdaptOrch: Task-Adaptive Multi-Agent Orchestration in the Era of LLM Performance Convergence
von: Yu, Geunbin
Veröffentlicht: (2026)
von: Yu, Geunbin
Veröffentlicht: (2026)
Territory Paint Wars: Diagnosing and Mitigating Failure Modes in Competitive Multi-Agent PPO
von: Singh, Diyansha
Veröffentlicht: (2026)
von: Singh, Diyansha
Veröffentlicht: (2026)
Fuzzy, Symbolic, and Contextual: Enhancing LLM Instruction via Cognitive Scaffolding
von: Figueiredo, Vanessa
Veröffentlicht: (2025)
von: Figueiredo, Vanessa
Veröffentlicht: (2025)
Your Data, My Model: Learning Who Really Helps in Federated Learning
von: Abdurakhmanova, Shamsiiat, et al.
Veröffentlicht: (2024)
von: Abdurakhmanova, Shamsiiat, et al.
Veröffentlicht: (2024)
When Actions Disappear: Adversarial Action Removal in Self-Play Reinforcement Learning
von: Kujur, Arahan
Veröffentlicht: (2026)
von: Kujur, Arahan
Veröffentlicht: (2026)
Dynamic Dual-Granularity Skill Bank for Agentic RL
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
von: Tu, Songjun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents
von: Tang, Wenjie, et al.
Veröffentlicht: (2026) -
Analysing Factorizations of Action-Value Networks for Cooperative Multi-Agent Reinforcement Learning
von: Castellini, Jacopo, et al.
Veröffentlicht: (2019) -
ME-IGM: Individual-Global-Max in Maximum Entropy Multi-Agent Reinforcement Learning
von: Chen, Wen-Tse, et al.
Veröffentlicht: (2024) -
Difference Rewards Policy Gradients
von: Castellini, Jacopo, et al.
Veröffentlicht: (2020) -
Safe and Policy-Compliant Multi-Agent Orchestration for Enterprise AI
von: Pasupuleti, Vinil, et al.
Veröffentlicht: (2026)