Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Shuo, Chen, Tianle, Amiri, Ryan, Amato, Christopher |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Usable Agent Discovery for Decentralized AI Systems
por: Dazzi, Patrizio, et al.
Publicado: (2026)
por: Dazzi, Patrizio, et al.
Publicado: (2026)
EvoGit: Decentralized Code Evolution via Git-Based Multi-Agent Collaboration
por: Huang, Beichen, et al.
Publicado: (2025)
por: Huang, Beichen, et al.
Publicado: (2025)
Optimizing PyTorch Inference with LLM-Based Multi-Agent Systems
por: Nagaitsev, Kirill, et al.
Publicado: (2025)
por: Nagaitsev, Kirill, et al.
Publicado: (2025)
SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks
por: Jose, Edwin
Publicado: (2026)
por: Jose, Edwin
Publicado: (2026)
DynTaskMAS: A Dynamic Task Graph-driven Framework for Asynchronous and Parallel LLM-based Multi-Agent Systems
por: Yu, Junwei, et al.
Publicado: (2025)
por: Yu, Junwei, et al.
Publicado: (2025)
Multi-Agent eXperimenter (MAX)
por: Gürcan, Önder
Publicado: (2024)
por: Gürcan, Önder
Publicado: (2024)
HEnRY: A Multi-Agent System Framework for Multi-Domain Contexts
por: Lacavalla, Emmanuele, et al.
Publicado: (2024)
por: Lacavalla, Emmanuele, et al.
Publicado: (2024)
PBFT-Backed Semantic Voting for Multi-Agent Memory Pruning
por: Bach, Duong
Publicado: (2025)
por: Bach, Duong
Publicado: (2025)
Communication-Efficient Training Workload Balancing for Decentralized Multi-Agent Learning
por: Mohammadabadi, Seyed Mahmoud Sajjadi, et al.
Publicado: (2024)
por: Mohammadabadi, Seyed Mahmoud Sajjadi, et al.
Publicado: (2024)
A biologically Inspired Trust Model for Open Multi-Agent Systems that is Resilient to Rapid Performance Fluctuations
por: Lygizou, Zoi, et al.
Publicado: (2025)
por: Lygizou, Zoi, et al.
Publicado: (2025)
Towards Enterprise-Ready Computer Using Generalist Agent
por: Marreed, Sami, et al.
Publicado: (2025)
por: Marreed, Sami, et al.
Publicado: (2025)
ITBench: Evaluating AI Agents across Diverse Real-World IT Automation Tasks
por: Jha, Saurabh, et al.
Publicado: (2025)
por: Jha, Saurabh, et al.
Publicado: (2025)
Efficient Tree-Structured Deep Research with Adaptive Resource Allocation
por: Nie, Lunyiu, et al.
Publicado: (2025)
por: Nie, Lunyiu, et al.
Publicado: (2025)
LAFA: Agentic LLM-Driven Federated Analytics over Decentralized Data Sources
por: Ji, Haichao, et al.
Publicado: (2025)
por: Ji, Haichao, et al.
Publicado: (2025)
AGNT2: Autonomous Agent Economies on Interaction-Optimized Layer 2 Infrastructure
por: Ruan, Anbang, et al.
Publicado: (2026)
por: Ruan, Anbang, et al.
Publicado: (2026)
S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination
por: Khan, Sajjad
Publicado: (2026)
por: Khan, Sajjad
Publicado: (2026)
Fully Distributed Fog Load Balancing with Multi-Agent Reinforcement Learning
por: Ebrahim, Maad, et al.
Publicado: (2024)
por: Ebrahim, Maad, et al.
Publicado: (2024)
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
por: Pan, Zaifeng, et al.
Publicado: (2025)
por: Pan, Zaifeng, et al.
Publicado: (2025)
Aegis: Taxonomy and Optimizations for Overcoming Agent-Environment Failures in LLM Agents
por: Song, Kevin, et al.
Publicado: (2025)
por: Song, Kevin, et al.
Publicado: (2025)
Moving From Monolithic To Microservices Architecture for Multi-Agent Systems
por: Goyal, Muskaan, et al.
Publicado: (2025)
por: Goyal, Muskaan, et al.
Publicado: (2025)
Decentralized Federated Policy Gradient with Byzantine Fault-Tolerance and Provably Fast Convergence
por: Jordan, Philip, et al.
Publicado: (2024)
por: Jordan, Philip, et al.
Publicado: (2024)
Several Performance Bounds on Decentralized Online Optimization are Highly Conservative and Potentially Misleading
por: Meunier, Erwan, et al.
Publicado: (2025)
por: Meunier, Erwan, et al.
Publicado: (2025)
Collective Privacy Recovery: Data-sharing Coordination via Decentralized Artificial Intelligence
por: Pournaras, Evangelos, et al.
Publicado: (2023)
por: Pournaras, Evangelos, et al.
Publicado: (2023)
APWA: A Distributed Architecture for Parallelizable Agentic Workflows
por: Rose, Evan, et al.
Publicado: (2026)
por: Rose, Evan, et al.
Publicado: (2026)
VineLM: Trie-Based Fine-Grained Control for Agentic Workflows
por: Pagonas, Nikos, et al.
Publicado: (2026)
por: Pagonas, Nikos, et al.
Publicado: (2026)
Distributed Online Rollout for Multivehicle Routing in Unmapped Environments
por: Weber, Jamison W., et al.
Publicado: (2023)
por: Weber, Jamison W., et al.
Publicado: (2023)
Speculative Actions: A Lossless Framework for Faster Agentic Systems
por: Ye, Naimeng, et al.
Publicado: (2025)
por: Ye, Naimeng, et al.
Publicado: (2025)
Holonic Learning: A Flexible Agent-based Distributed Machine Learning Framework
por: Esmaeili, Ahmad, et al.
Publicado: (2023)
por: Esmaeili, Ahmad, et al.
Publicado: (2023)
Single-Loop Federated Actor-Critic across Heterogeneous Environments
por: Zhu, Ye, et al.
Publicado: (2024)
por: Zhu, Ye, et al.
Publicado: (2024)
Decoupling Correctness from Policy: A Deterministic Causal Structure for Multi-Agent Systems
por: Ren, Zhiyuan, et al.
Publicado: (2025)
por: Ren, Zhiyuan, et al.
Publicado: (2025)
AI Metropolis: Scaling Large Language Model-based Multi-Agent Simulation with Out-of-order Execution
por: Xie, Zhiqiang, et al.
Publicado: (2024)
por: Xie, Zhiqiang, et al.
Publicado: (2024)
MOD-X: A Modular Open Decentralized eXchange Framework proposal for Heterogeneous Interoperable Artificial Intelligence Agents
por: Ioannides, Georgios, et al.
Publicado: (2025)
por: Ioannides, Georgios, et al.
Publicado: (2025)
Empowering Scientific Workflows with Federated Agents
por: Kamatar, Alok, et al.
Publicado: (2025)
por: Kamatar, Alok, et al.
Publicado: (2025)
Incentivizing Multi-Tenant Split Federated Learning for Foundation Models at the Network Edge
por: Li, Songyuan, et al.
Publicado: (2025)
por: Li, Songyuan, et al.
Publicado: (2025)
Pythia: Exploiting Workflow Predictability for Efficient Agent-Native LLM Serving
por: Yu, Shan, et al.
Publicado: (2026)
por: Yu, Shan, et al.
Publicado: (2026)
Rudder: Steering Prefetching in Distributed GNN Training using LLM Agents
por: Sarkar, Aishwarya, et al.
Publicado: (2026)
por: Sarkar, Aishwarya, et al.
Publicado: (2026)
AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds
por: Chen, Yinfang, et al.
Publicado: (2025)
por: Chen, Yinfang, et al.
Publicado: (2025)
When Computing follows Vehicles: Decentralized Mobility-Aware Resource Allocation for Edge-to-Cloud Continuum
por: Nezami, Zeinab, et al.
Publicado: (2024)
por: Nezami, Zeinab, et al.
Publicado: (2024)
Ensuring Fair LLM Serving Amid Diverse Applications
por: Khan, Redwan Ibne Seraj, et al.
Publicado: (2024)
por: Khan, Redwan Ibne Seraj, et al.
Publicado: (2024)
Summary Paper: Use Case on Building Collaborative Safe Autonomous Systems-A Robotdog for Guiding Visually Impaired People
por: Malhotra, Aman, et al.
Publicado: (2024)
por: Malhotra, Aman, et al.
Publicado: (2024)
Ejemplares similares
-
Usable Agent Discovery for Decentralized AI Systems
por: Dazzi, Patrizio, et al.
Publicado: (2026) -
EvoGit: Decentralized Code Evolution via Git-Based Multi-Agent Collaboration
por: Huang, Beichen, et al.
Publicado: (2025) -
Optimizing PyTorch Inference with LLM-Based Multi-Agent Systems
por: Nagaitsev, Kirill, et al.
Publicado: (2025) -
SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks
por: Jose, Edwin
Publicado: (2026) -
DynTaskMAS: A Dynamic Task Graph-driven Framework for Asynchronous and Parallel LLM-based Multi-Agent Systems
por: Yu, Junwei, et al.
Publicado: (2025)