Optimizing PyTorch Inference with LLM-Based Multi-Agent Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Nagaitsev, Kirill, Grbcic, Luka, Williams, Samuel, Iancu, Costin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic
by: Liu, Shuo, et al.
Published: (2026)
by: Liu, Shuo, et al.
Published: (2026)
DynTaskMAS: A Dynamic Task Graph-driven Framework for Asynchronous and Parallel LLM-based Multi-Agent Systems
by: Yu, Junwei, et al.
Published: (2025)
by: Yu, Junwei, et al.
Published: (2025)
HEnRY: A Multi-Agent System Framework for Multi-Domain Contexts
by: Lacavalla, Emmanuele, et al.
Published: (2024)
by: Lacavalla, Emmanuele, et al.
Published: (2024)
A biologically Inspired Trust Model for Open Multi-Agent Systems that is Resilient to Rapid Performance Fluctuations
by: Lygizou, Zoi, et al.
Published: (2025)
by: Lygizou, Zoi, et al.
Published: (2025)
Multi-Agent eXperimenter (MAX)
by: Gürcan, Önder
Published: (2024)
by: Gürcan, Önder
Published: (2024)
Usable Agent Discovery for Decentralized AI Systems
by: Dazzi, Patrizio, et al.
Published: (2026)
by: Dazzi, Patrizio, et al.
Published: (2026)
PBFT-Backed Semantic Voting for Multi-Agent Memory Pruning
by: Bach, Duong
Published: (2025)
by: Bach, Duong
Published: (2025)
AGNT2: Autonomous Agent Economies on Interaction-Optimized Layer 2 Infrastructure
by: Ruan, Anbang, et al.
Published: (2026)
by: Ruan, Anbang, et al.
Published: (2026)
SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks
by: Jose, Edwin
Published: (2026)
by: Jose, Edwin
Published: (2026)
Towards Enterprise-Ready Computer Using Generalist Agent
by: Marreed, Sami, et al.
Published: (2025)
by: Marreed, Sami, et al.
Published: (2025)
S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination
by: Khan, Sajjad
Published: (2026)
by: Khan, Sajjad
Published: (2026)
ITBench: Evaluating AI Agents across Diverse Real-World IT Automation Tasks
by: Jha, Saurabh, et al.
Published: (2025)
by: Jha, Saurabh, et al.
Published: (2025)
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
by: Pan, Zaifeng, et al.
Published: (2025)
by: Pan, Zaifeng, et al.
Published: (2025)
Aegis: Taxonomy and Optimizations for Overcoming Agent-Environment Failures in LLM Agents
by: Song, Kevin, et al.
Published: (2025)
by: Song, Kevin, et al.
Published: (2025)
Moving From Monolithic To Microservices Architecture for Multi-Agent Systems
by: Goyal, Muskaan, et al.
Published: (2025)
by: Goyal, Muskaan, et al.
Published: (2025)
Speculative Actions: A Lossless Framework for Faster Agentic Systems
by: Ye, Naimeng, et al.
Published: (2025)
by: Ye, Naimeng, et al.
Published: (2025)
VineLM: Trie-Based Fine-Grained Control for Agentic Workflows
by: Pagonas, Nikos, et al.
Published: (2026)
by: Pagonas, Nikos, et al.
Published: (2026)
torch-sla: Differentiable Sparse Linear Algebra with Adjoint Solvers and Sparse Tensor Parallelism for PyTorch
by: Chi, Mingyuan, et al.
Published: (2026)
by: Chi, Mingyuan, et al.
Published: (2026)
Fully Distributed Fog Load Balancing with Multi-Agent Reinforcement Learning
by: Ebrahim, Maad, et al.
Published: (2024)
by: Ebrahim, Maad, et al.
Published: (2024)
Distributed Online Rollout for Multivehicle Routing in Unmapped Environments
by: Weber, Jamison W., et al.
Published: (2023)
by: Weber, Jamison W., et al.
Published: (2023)
APWA: A Distributed Architecture for Parallelizable Agentic Workflows
by: Rose, Evan, et al.
Published: (2026)
by: Rose, Evan, et al.
Published: (2026)
Efficient Tree-Structured Deep Research with Adaptive Resource Allocation
by: Nie, Lunyiu, et al.
Published: (2025)
by: Nie, Lunyiu, et al.
Published: (2025)
Decoupling Correctness from Policy: A Deterministic Causal Structure for Multi-Agent Systems
by: Ren, Zhiyuan, et al.
Published: (2025)
by: Ren, Zhiyuan, et al.
Published: (2025)
AI Metropolis: Scaling Large Language Model-based Multi-Agent Simulation with Out-of-order Execution
by: Xie, Zhiqiang, et al.
Published: (2024)
by: Xie, Zhiqiang, et al.
Published: (2024)
EvoGit: Decentralized Code Evolution via Git-Based Multi-Agent Collaboration
by: Huang, Beichen, et al.
Published: (2025)
by: Huang, Beichen, et al.
Published: (2025)
TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training
by: Liang, Wanchao, et al.
Published: (2024)
by: Liang, Wanchao, et al.
Published: (2024)
Rudder: Steering Prefetching in Distributed GNN Training using LLM Agents
by: Sarkar, Aishwarya, et al.
Published: (2026)
by: Sarkar, Aishwarya, et al.
Published: (2026)
ASIC-Agent: An Autonomous Multi-Agent System for ASIC Design with Benchmark Evaluation
by: Allam, Ahmed, et al.
Published: (2025)
by: Allam, Ahmed, et al.
Published: (2025)
Ensuring Fair LLM Serving Amid Diverse Applications
by: Khan, Redwan Ibne Seraj, et al.
Published: (2024)
by: Khan, Redwan Ibne Seraj, et al.
Published: (2024)
Self-Supervised Inference of Agents in Trustless Environments
by: Larin, Vladyslav, et al.
Published: (2024)
by: Larin, Vladyslav, et al.
Published: (2024)
Communication-Efficient Training Workload Balancing for Decentralized Multi-Agent Learning
by: Mohammadabadi, Seyed Mahmoud Sajjadi, et al.
Published: (2024)
by: Mohammadabadi, Seyed Mahmoud Sajjadi, et al.
Published: (2024)
LAFA: Agentic LLM-Driven Federated Analytics over Decentralized Data Sources
by: Ji, Haichao, et al.
Published: (2025)
by: Ji, Haichao, et al.
Published: (2025)
Holonic Learning: A Flexible Agent-based Distributed Machine Learning Framework
by: Esmaeili, Ahmad, et al.
Published: (2023)
by: Esmaeili, Ahmad, et al.
Published: (2023)
AIOpsLab: A Holistic Framework to Evaluate AI Agents for Enabling Autonomous Clouds
by: Chen, Yinfang, et al.
Published: (2025)
by: Chen, Yinfang, et al.
Published: (2025)
Warp-Cortex: An Asynchronous, Memory-Efficient Architecture for Million-Agent Cognitive Scaling on Consumer Hardware
by: Williams, Jorge L. Ruiz
Published: (2026)
by: Williams, Jorge L. Ruiz
Published: (2026)
Towards Blockchain-based Multi-Agent Robotic Systems: Analysis, Classification and Applications
by: Afanasyev, Ilya, et al.
Published: (2019)
by: Afanasyev, Ilya, et al.
Published: (2019)
Several Performance Bounds on Decentralized Online Optimization are Highly Conservative and Potentially Misleading
by: Meunier, Erwan, et al.
Published: (2025)
by: Meunier, Erwan, et al.
Published: (2025)
When Agents Control Robots: A Zero Trust Policy Model for Agentic Cyber-Physical Systems
by: Ranathunga, Tharindu, et al.
Published: (2026)
by: Ranathunga, Tharindu, et al.
Published: (2026)
Agent-Based Triangle Counting: Unlocking Truss Decomposition, Triangle Centrality, and Local Clustering Coefficient
by: Chand, Prabhat Kumar, et al.
Published: (2024)
by: Chand, Prabhat Kumar, et al.
Published: (2024)
Characterization and Mitigation of Insufficiencies in Automated Driving Systems
by: Fu, Yuting, et al.
Published: (2024)
by: Fu, Yuting, et al.
Published: (2024)
Similar Items
-
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic
by: Liu, Shuo, et al.
Published: (2026) -
DynTaskMAS: A Dynamic Task Graph-driven Framework for Asynchronous and Parallel LLM-based Multi-Agent Systems
by: Yu, Junwei, et al.
Published: (2025) -
HEnRY: A Multi-Agent System Framework for Multi-Domain Contexts
by: Lacavalla, Emmanuele, et al.
Published: (2024) -
A biologically Inspired Trust Model for Open Multi-Agent Systems that is Resilient to Rapid Performance Fluctuations
by: Lygizou, Zoi, et al.
Published: (2025) -
Multi-Agent eXperimenter (MAX)
by: Gürcan, Önder
Published: (2024)