Connecting Large Language Model Agent to High Performance Computing Resource
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Heng, Brace, Alexander, Siebenschuh, Carlo, Pauloski, Greg, Foster, Ian, Ramanathan, Arvind |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SparkAttention: High-Performance Multi-Head Attention for Large Models on Volta GPU Architecture
von: Xu, Youxuan, et al.
Veröffentlicht: (2025)
von: Xu, Youxuan, et al.
Veröffentlicht: (2025)
Comparison of Autoscaling Frameworks for Containerised Machine-Learning-Applications in a Local and Cloud Environment
von: Schroeder, Christian, et al.
Veröffentlicht: (2023)
von: Schroeder, Christian, et al.
Veröffentlicht: (2023)
CodeCRDT: Observation-Driven Coordination for Multi-Agent LLM Code Generation
von: Pugachev, Sergey
Veröffentlicht: (2025)
von: Pugachev, Sergey
Veröffentlicht: (2025)
Aergia: Leveraging Heterogeneity in Federated Learning Systems
von: Cox, Bart, et al.
Veröffentlicht: (2022)
von: Cox, Bart, et al.
Veröffentlicht: (2022)
Roadmap for Edge AI: A Dagstuhl Perspective
von: Ding, Aaron Yi, et al.
Veröffentlicht: (2021)
von: Ding, Aaron Yi, et al.
Veröffentlicht: (2021)
Towards Optimal Heterogeneous Client Sampling in Multi-Model Federated Learning
von: Zhang, Haoran, et al.
Veröffentlicht: (2025)
von: Zhang, Haoran, et al.
Veröffentlicht: (2025)
Parameterizing Federated Continual Learning for Reproducible Research
von: Cox, Bart, et al.
Veröffentlicht: (2024)
von: Cox, Bart, et al.
Veröffentlicht: (2024)
Asynchronous Byzantine Federated Learning
von: Cox, Bart, et al.
Veröffentlicht: (2024)
von: Cox, Bart, et al.
Veröffentlicht: (2024)
Hyper-parameter Optimization for Federated Learning with Step-wise Adaptive Mechanism
von: Saadati, Yasaman, et al.
Veröffentlicht: (2024)
von: Saadati, Yasaman, et al.
Veröffentlicht: (2024)
Training Diffusion Models with Federated Learning
von: de Goede, Matthijs, et al.
Veröffentlicht: (2024)
von: de Goede, Matthijs, et al.
Veröffentlicht: (2024)
Quantize Once, Train Fast: Allreduce-Compatible Compression with Provable Guarantees
von: Xin, Jihao, et al.
Veröffentlicht: (2023)
von: Xin, Jihao, et al.
Veröffentlicht: (2023)
Asynchronous Multi-Server Federated Learning for Geo-Distributed Clients
von: Zuo, Yuncong, et al.
Veröffentlicht: (2024)
von: Zuo, Yuncong, et al.
Veröffentlicht: (2024)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025)
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025)
SI-ChainFL: Shapley-Incentivized Secure Federated Learning for High-Speed Rail Data Sharing
von: Zhao, Mingjie, et al.
Veröffentlicht: (2026)
von: Zhao, Mingjie, et al.
Veröffentlicht: (2026)
DAGER: Exact Gradient Inversion for Large Language Models
von: Petrov, Ivo, et al.
Veröffentlicht: (2024)
von: Petrov, Ivo, et al.
Veröffentlicht: (2024)
Separating Intelligence from Execution: A Workflow Engine for the Model Context Protocol
von: Parmar, Abhinav Singh
Veröffentlicht: (2026)
von: Parmar, Abhinav Singh
Veröffentlicht: (2026)
FedPLT: Scalable, Resource-Efficient, and Heterogeneity-Aware Federated Learning via Partial Layer Training
von: Dabaja, Ahmad, et al.
Veröffentlicht: (2026)
von: Dabaja, Ahmad, et al.
Veröffentlicht: (2026)
Uncertainty Estimation in Multi-Agent Distributed Learning for AI-Enabled Edge Devices
von: Radchenko, Gleb, et al.
Veröffentlicht: (2024)
von: Radchenko, Gleb, et al.
Veröffentlicht: (2024)
AutoDDL: Automatic Distributed Deep Learning with Near-Optimal Bandwidth Cost
von: Chen, Jinfan, et al.
Veröffentlicht: (2023)
von: Chen, Jinfan, et al.
Veröffentlicht: (2023)
WLB-LLM: Workload-Balanced 4D Parallelism for Large Language Model Training
von: Wang, Zheng, et al.
Veröffentlicht: (2025)
von: Wang, Zheng, et al.
Veröffentlicht: (2025)
SPEAR++: Scaling Gradient Inversion via Sparsely-Used Dictionary Learning
von: Bakarsky, Alexander, et al.
Veröffentlicht: (2025)
von: Bakarsky, Alexander, et al.
Veröffentlicht: (2025)
TokenCake: A KV-Cache-centric Serving Framework for LLM-based Multi-Agent Applications
von: Bian, Zhuohang, et al.
Veröffentlicht: (2025)
von: Bian, Zhuohang, et al.
Veröffentlicht: (2025)
Libra: Unleashing GPU Heterogeneity for High-Performance Sparse Matrix Multiplication
von: Shi, Jinliang, et al.
Veröffentlicht: (2025)
von: Shi, Jinliang, et al.
Veröffentlicht: (2025)
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
ADF-LoRA: Alternating Low-Rank Aggregation for Decentralized Federated Fine-Tuning
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2025)
FlashSparse: Minimizing Computation Redundancy for Fast Sparse Matrix Multiplications on Tensor Cores
von: Shi, Jinliang, et al.
Veröffentlicht: (2024)
von: Shi, Jinliang, et al.
Veröffentlicht: (2024)
Combinatorial Client-Master Multiagent Deep Reinforcement Learning for Task Offloading in Mobile Edge Computing
von: Gebrekidan, Tesfay Zemuy, et al.
Veröffentlicht: (2024)
von: Gebrekidan, Tesfay Zemuy, et al.
Veröffentlicht: (2024)
Serial Parallel Reliability Redundancy Allocation Optimization for Energy Efficient and Fault Tolerant Cloud Computing
von: Krishna, Gutha Jaya
Veröffentlicht: (2024)
von: Krishna, Gutha Jaya
Veröffentlicht: (2024)
Chimera: Efficiently Training Large-Scale Neural Networks with Bidirectional Pipelines
von: Li, Shigang, et al.
Veröffentlicht: (2021)
von: Li, Shigang, et al.
Veröffentlicht: (2021)
TPI-LLM: Serving 70B-scale LLMs Efficiently on Low-resource Edge Devices
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
Edge AI Collaborative Learning: Bayesian Approaches to Uncertainty Estimation
von: Radchenko, Gleb, et al.
Veröffentlicht: (2024)
von: Radchenko, Gleb, et al.
Veröffentlicht: (2024)
A Full Compression Pipeline for Green Federated Learning in Communication-Constrained Environments
von: Colybes, Elouan, et al.
Veröffentlicht: (2026)
von: Colybes, Elouan, et al.
Veröffentlicht: (2026)
Learning In Chaos: Efficient Autoscaling and Self-Healing for Multi-Party Distributed Training
von: Feng, Wenjiao, et al.
Veröffentlicht: (2025)
von: Feng, Wenjiao, et al.
Veröffentlicht: (2025)
SPEAR:Exact Gradient Inversion of Batches in Federated Learning
von: Dimitrov, Dimitar I., et al.
Veröffentlicht: (2024)
von: Dimitrov, Dimitar I., et al.
Veröffentlicht: (2024)
Federated Learning with Differential Privacy
von: Banse, Adrien, et al.
Veröffentlicht: (2024)
von: Banse, Adrien, et al.
Veröffentlicht: (2024)
Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters
von: Li, Zonghang, et al.
Veröffentlicht: (2025)
von: Li, Zonghang, et al.
Veröffentlicht: (2025)
Adaptive GPU Resource Allocation for Multi-Agent Collaborative Reasoning in Serverless Environments
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
Towards Building Private LLMs: Exploring Multi-Node Expert Parallelism on Apple Silicon for Mixture-of-Experts Large Language Model
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2025)
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2025)
Robustness of deep learning classification to adversarial input on GPUs: asynchronous parallel accumulation is a source of vulnerability
von: Shanmugavelu, Sanjif, et al.
Veröffentlicht: (2025)
von: Shanmugavelu, Sanjif, et al.
Veröffentlicht: (2025)
GRAIN: Exact Graph Reconstruction from Gradients
von: Drencheva, Maria, et al.
Veröffentlicht: (2025)
von: Drencheva, Maria, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SparkAttention: High-Performance Multi-Head Attention for Large Models on Volta GPU Architecture
von: Xu, Youxuan, et al.
Veröffentlicht: (2025) -
Comparison of Autoscaling Frameworks for Containerised Machine-Learning-Applications in a Local and Cloud Environment
von: Schroeder, Christian, et al.
Veröffentlicht: (2023) -
CodeCRDT: Observation-Driven Coordination for Multi-Agent LLM Code Generation
von: Pugachev, Sergey
Veröffentlicht: (2025) -
Aergia: Leveraging Heterogeneity in Federated Learning Systems
von: Cox, Bart, et al.
Veröffentlicht: (2022) -
Roadmap for Edge AI: A Dagstuhl Perspective
von: Ding, Aaron Yi, et al.
Veröffentlicht: (2021)