Adaptive GPU Resource Allocation for Multi-Agent Collaborative Reasoning in Serverless Environments
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Guilin, Guo, Wulan, Tan, Ziqi |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
par: Zhang, Guilin, et autres
Publié: (2025)
par: Zhang, Guilin, et autres
Publié: (2025)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
par: Murimi, Almond Kiruthu
Publié: (2025)
par: Murimi, Almond Kiruthu
Publié: (2025)
Agent Identity URI Scheme: Topology-Independent Naming and Capability-Based Discovery for Multi-Agent Systems
par: Rodriguez Jr, Roland R.
Publié: (2026)
par: Rodriguez Jr, Roland R.
Publié: (2026)
The Bureaucracy of Speed: Structural Equivalence Between Memory Consistency Models and Multi-Agent Authorization Revocation
par: Parakhin, Vladyslav
Publié: (2026)
par: Parakhin, Vladyslav
Publié: (2026)
Accelerating Geo-distributed Machine Learning with Network-Aware Adaptive Tree and Auxiliary Route
par: Li, Zonghang, et autres
Publié: (2024)
par: Li, Zonghang, et autres
Publié: (2024)
Token Coherence: Adapting MESI Cache Protocols to Minimize Synchronization Overhead in Multi-Agent LLM Systems
par: Parakhin, Vladyslav
Publié: (2026)
par: Parakhin, Vladyslav
Publié: (2026)
HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean Aggregation
par: Nguyen, Thinh, et autres
Publié: (2025)
par: Nguyen, Thinh, et autres
Publié: (2025)
Edge AI Collaborative Learning: Bayesian Approaches to Uncertainty Estimation
par: Radchenko, Gleb, et autres
Publié: (2024)
par: Radchenko, Gleb, et autres
Publié: (2024)
From Logic Monopoly to Social Contract: Separation of Power and the Institutional Foundations for Autonomous Agent Economies
par: Ruan, Anbang
Publié: (2026)
par: Ruan, Anbang
Publié: (2026)
Aethon: A Reference-Based Replication Primitive for Constant-Time Instantiation of Stateful AI Agents
par: Rao, Swanand, et autres
Publié: (2026)
par: Rao, Swanand, et autres
Publié: (2026)
SparkAttention: High-Performance Multi-Head Attention for Large Models on Volta GPU Architecture
par: Xu, Youxuan, et autres
Publié: (2025)
par: Xu, Youxuan, et autres
Publié: (2025)
Impact of Network Topology on Byzantine Resilience in Decentralized Federated Learning
par: Bhattacharya, Siddhartha, et autres
Publié: (2024)
par: Bhattacharya, Siddhartha, et autres
Publié: (2024)
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
par: Polyakov, Igor, et autres
Publié: (2025)
par: Polyakov, Igor, et autres
Publié: (2025)
TokenCake: A KV-Cache-centric Serving Framework for LLM-based Multi-Agent Applications
par: Bian, Zhuohang, et autres
Publié: (2025)
par: Bian, Zhuohang, et autres
Publié: (2025)
Federated Learning Model Aggregation in Heterogenous Aerial and Space Networks
par: Dong, Fan, et autres
Publié: (2023)
par: Dong, Fan, et autres
Publié: (2023)
Scheduling the Unschedulable: Taming Black-Box LLM Inference at Scale
par: Yuan, Renzhong, et autres
Publié: (2026)
par: Yuan, Renzhong, et autres
Publié: (2026)
Learning to Collaborate: An Orchestrated-Decentralized Framework for Peer-to-Peer LLM Federation
par: Singh, Inderjeet, et autres
Publié: (2026)
par: Singh, Inderjeet, et autres
Publié: (2026)
A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning
par: Rosendal, Daan, et autres
Publié: (2026)
par: Rosendal, Daan, et autres
Publié: (2026)
Connecting Large Language Model Agent to High Performance Computing Resource
par: Ma, Heng, et autres
Publié: (2025)
par: Ma, Heng, et autres
Publié: (2025)
Bridging Generalization Gap of Heterogeneous Federated Clients Using Generative Models
par: Niu, Ziru, et autres
Publié: (2025)
par: Niu, Ziru, et autres
Publié: (2025)
Serverless GPU Architecture for Enterprise HR Analytics: A Production-Scale BDaaS Implementation
par: Zhang, Guilin, et autres
Publié: (2025)
par: Zhang, Guilin, et autres
Publié: (2025)
FLEdge: Benchmarking Federated Machine Learning Applications in Edge Computing Systems
par: Woisetschläger, Herbert, et autres
Publié: (2023)
par: Woisetschläger, Herbert, et autres
Publié: (2023)
CarbonEdge: Carbon-Aware Deep Learning Inference Framework for Sustainable Edge Computing
par: Zhang, Guilin, et autres
Publié: (2026)
par: Zhang, Guilin, et autres
Publié: (2026)
AutoDDL: Automatic Distributed Deep Learning with Near-Optimal Bandwidth Cost
par: Chen, Jinfan, et autres
Publié: (2023)
par: Chen, Jinfan, et autres
Publié: (2023)
CodeCRDT: Observation-Driven Coordination for Multi-Agent LLM Code Generation
par: Pugachev, Sergey
Publié: (2025)
par: Pugachev, Sergey
Publié: (2025)
Context Engineering: From Prompts to Corporate Multi-Agent Architecture
par: Vishnyakova, Vera V.
Publié: (2026)
par: Vishnyakova, Vera V.
Publié: (2026)
Complex Event Processing in the Edge: A Combined Optimization Approach for Data and Code Placement
par: Uyanık, Halit, et autres
Publié: (2026)
par: Uyanık, Halit, et autres
Publié: (2026)
FedStrategist: A Meta-Learning Framework for Adaptive and Robust Aggregation in Federated Learning
par: Haque, Md Rafid, et autres
Publié: (2025)
par: Haque, Md Rafid, et autres
Publié: (2025)
Federated Few-Shot Learning on Neuromorphic Hardware: An Empirical Study Across Physical Edge Nodes
par: Motta, Steven, et autres
Publié: (2026)
par: Motta, Steven, et autres
Publié: (2026)
Comparison of Autoscaling Frameworks for Containerised Machine-Learning-Applications in a Local and Cloud Environment
par: Schroeder, Christian, et autres
Publié: (2023)
par: Schroeder, Christian, et autres
Publié: (2023)
Federated Fine-Tuning of LLMs on the Very Edge: The Good, the Bad, the Ugly
par: Woisetschläger, Herbert, et autres
Publié: (2023)
par: Woisetschläger, Herbert, et autres
Publié: (2023)
Hyper-parameter Optimization for Federated Learning with Step-wise Adaptive Mechanism
par: Saadati, Yasaman, et autres
Publié: (2024)
par: Saadati, Yasaman, et autres
Publié: (2024)
Mobile Traffic Prediction at the Edge Through Distributed and Deep Transfer Learning
par: Petrella, Alfredo, et autres
Publié: (2023)
par: Petrella, Alfredo, et autres
Publié: (2023)
ChargingBoul: A Competitive Negotiating Agent with Novel Opponent Modeling
par: Shymanski, Joe
Publié: (2025)
par: Shymanski, Joe
Publié: (2025)
Libra: Unleashing GPU Heterogeneity for High-Performance Sparse Matrix Multiplication
par: Shi, Jinliang, et autres
Publié: (2025)
par: Shi, Jinliang, et autres
Publié: (2025)
Asynchronous Multi-Server Federated Learning for Geo-Distributed Clients
par: Zuo, Yuncong, et autres
Publié: (2024)
par: Zuo, Yuncong, et autres
Publié: (2024)
Towards Optimal Heterogeneous Client Sampling in Multi-Model Federated Learning
par: Zhang, Haoran, et autres
Publié: (2025)
par: Zhang, Haoran, et autres
Publié: (2025)
Decentralized Task Scheduling in Distributed Systems: A Deep Reinforcement Learning Approach
par: John, Daniel Benniah
Publié: (2026)
par: John, Daniel Benniah
Publié: (2026)
A Framework for testing Federated Learning algorithms using an edge-like environment
par: Schwanck, Felipe Machado, et autres
Publié: (2024)
par: Schwanck, Felipe Machado, et autres
Publié: (2024)
HECATE: An ECS-based Framework for Teaching and Developing Multi-Agent Systems
par: Casals, Arthur, et autres
Publié: (2025)
par: Casals, Arthur, et autres
Publié: (2025)
Documents similaires
-
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
par: Zhang, Guilin, et autres
Publié: (2025) -
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
par: Murimi, Almond Kiruthu
Publié: (2025) -
Agent Identity URI Scheme: Topology-Independent Naming and Capability-Based Discovery for Multi-Agent Systems
par: Rodriguez Jr, Roland R.
Publié: (2026) -
The Bureaucracy of Speed: Structural Equivalence Between Memory Consistency Models and Multi-Agent Authorization Revocation
par: Parakhin, Vladyslav
Publié: (2026) -
Accelerating Geo-distributed Machine Learning with Network-Aware Adaptive Tree and Auxiliary Route
par: Li, Zonghang, et autres
Publié: (2024)