Federated Fine-Tuning of LLMs on the Very Edge: The Good, the Bad, the Ugly
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Woisetschläger, Herbert, Isenko, Alexander, Wang, Shiqiang, Mayer, Ruben, Jacobsen, Hans-Arno |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FLEdge: Benchmarking Federated Machine Learning Applications in Edge Computing Systems
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2023)
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2023)
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
von: Erben, Alexander, et al.
Veröffentlicht: (2023)
von: Erben, Alexander, et al.
Veröffentlicht: (2023)
A Survey on Efficient Federated Learning Methods for Foundation Model Training
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2024)
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2024)
Scheduling the Unschedulable: Taming Black-Box LLM Inference at Scale
von: Yuan, Renzhong, et al.
Veröffentlicht: (2026)
von: Yuan, Renzhong, et al.
Veröffentlicht: (2026)
Designing Scalable Rate Limiting Systems: Algorithms, Architecture, and Distributed Solutions
von: Guan, Bo
Veröffentlicht: (2026)
von: Guan, Bo
Veröffentlicht: (2026)
Federated Learning and AI Regulation in the European Union: Who is Responsible? -- An Interdisciplinary Analysis
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2024)
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2024)
Impact of Network Topology on Byzantine Resilience in Decentralized Federated Learning
von: Bhattacharya, Siddhartha, et al.
Veröffentlicht: (2024)
von: Bhattacharya, Siddhartha, et al.
Veröffentlicht: (2024)
Cost-Aware Logging: Measuring the Financial Impact of Excessive Log Retention in Small-Scale Cloud Deployments
von: Putra, Jody Almaida
Veröffentlicht: (2026)
von: Putra, Jody Almaida
Veröffentlicht: (2026)
Libra: Unleashing GPU Heterogeneity for High-Performance Sparse Matrix Multiplication
von: Shi, Jinliang, et al.
Veröffentlicht: (2025)
von: Shi, Jinliang, et al.
Veröffentlicht: (2025)
Efficient Construction of Large Search Spaces for Auto-Tuning
von: Willemsen, Floris-Jan, et al.
Veröffentlicht: (2025)
von: Willemsen, Floris-Jan, et al.
Veröffentlicht: (2025)
How Machine Learning-Data Driven Replication Strategies Enhance Fault Tolerance in Large-Scale Distributed Systems
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025)
von: Murimi, Almond Kiruthu
Veröffentlicht: (2025)
TokenCake: A KV-Cache-centric Serving Framework for LLM-based Multi-Agent Applications
von: Bian, Zhuohang, et al.
Veröffentlicht: (2025)
von: Bian, Zhuohang, et al.
Veröffentlicht: (2025)
Complex Event Processing in the Edge: A Combined Optimization Approach for Data and Code Placement
von: Uyanık, Halit, et al.
Veröffentlicht: (2026)
von: Uyanık, Halit, et al.
Veröffentlicht: (2026)
AMP4EC: Adaptive Model Partitioning Framework for Efficient Deep Learning Inference in Edge Computing Environments
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
HFedATM: Hierarchical Federated Domain Generalization via Optimal Transport and Regularized Mean Aggregation
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
SparkAttention: High-Performance Multi-Head Attention for Large Models on Volta GPU Architecture
von: Xu, Youxuan, et al.
Veröffentlicht: (2025)
von: Xu, Youxuan, et al.
Veröffentlicht: (2025)
AutoDDL: Automatic Distributed Deep Learning with Near-Optimal Bandwidth Cost
von: Chen, Jinfan, et al.
Veröffentlicht: (2023)
von: Chen, Jinfan, et al.
Veröffentlicht: (2023)
Accelerating Geo-distributed Machine Learning with Network-Aware Adaptive Tree and Auxiliary Route
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
von: Li, Zonghang, et al.
Veröffentlicht: (2024)
Federated Learning Priorities Under the European Union Artificial Intelligence Act
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2024)
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2024)
Federated Learning Model Aggregation in Heterogenous Aerial and Space Networks
von: Dong, Fan, et al.
Veröffentlicht: (2023)
von: Dong, Fan, et al.
Veröffentlicht: (2023)
Training LLMs on HPC Systems: Best Practices from the OpenGPT-X Project
von: Penke, Carolin, et al.
Veröffentlicht: (2025)
von: Penke, Carolin, et al.
Veröffentlicht: (2025)
Mobile Traffic Prediction at the Edge Through Distributed and Deep Transfer Learning
von: Petrella, Alfredo, et al.
Veröffentlicht: (2023)
von: Petrella, Alfredo, et al.
Veröffentlicht: (2023)
Federated Few-Shot Learning on Neuromorphic Hardware: An Empirical Study Across Physical Edge Nodes
von: Motta, Steven, et al.
Veröffentlicht: (2026)
von: Motta, Steven, et al.
Veröffentlicht: (2026)
Bridging Generalization Gap of Heterogeneous Federated Clients Using Generative Models
von: Niu, Ziru, et al.
Veröffentlicht: (2025)
von: Niu, Ziru, et al.
Veröffentlicht: (2025)
Intent-driven scheduling of backup jobs
von: Dutta, Souvik, et al.
Veröffentlicht: (2024)
von: Dutta, Souvik, et al.
Veröffentlicht: (2024)
Deadline-Aware Joint Task Scheduling and Offloading in Mobile Edge Computing Systems
von: Nguyen, Ngoc Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Ngoc Hung, et al.
Veröffentlicht: (2025)
From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs
von: Kang, Daemyung, et al.
Veröffentlicht: (2026)
von: Kang, Daemyung, et al.
Veröffentlicht: (2026)
TAGC: Optimizing Gradient Communication in Distributed Transformer Training
von: Polyakov, Igor, et al.
Veröffentlicht: (2025)
von: Polyakov, Igor, et al.
Veröffentlicht: (2025)
Towards Building Private LLMs: Exploring Multi-Node Expert Parallelism on Apple Silicon for Mixture-of-Experts Large Language Model
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2025)
von: Chen, Mu-Chi, et al.
Veröffentlicht: (2025)
Near-Optimal Sparse Allreduce for Distributed Deep Learning
von: Li, Shigang, et al.
Veröffentlicht: (2022)
von: Li, Shigang, et al.
Veröffentlicht: (2022)
FlashSparse: Minimizing Computation Redundancy for Fast Sparse Matrix Multiplications on Tensor Cores
von: Shi, Jinliang, et al.
Veröffentlicht: (2024)
von: Shi, Jinliang, et al.
Veröffentlicht: (2024)
Chimera: Efficiently Training Large-Scale Neural Networks with Bidirectional Pipelines
von: Li, Shigang, et al.
Veröffentlicht: (2021)
von: Li, Shigang, et al.
Veröffentlicht: (2021)
Adaptive GPU Resource Allocation for Multi-Agent Collaborative Reasoning in Serverless Environments
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
A Taxonomy and Resolution Strategy for Client-Level Disagreements in Federated Learning
von: Rosendal, Daan, et al.
Veröffentlicht: (2026)
von: Rosendal, Daan, et al.
Veröffentlicht: (2026)
Heuristic Search Space Partitioning for Low-Latency Multi-Tenant Cloud Queries
von: Pathak, Prashant Kumar, et al.
Veröffentlicht: (2026)
von: Pathak, Prashant Kumar, et al.
Veröffentlicht: (2026)
Safety-Critical Edge Robotics Architecture with Bounded End-to-End Latency
von: Gala, Gautam, et al.
Veröffentlicht: (2024)
von: Gala, Gautam, et al.
Veröffentlicht: (2024)
Token Coherence: Adapting MESI Cache Protocols to Minimize Synchronization Overhead in Multi-Agent LLM Systems
von: Parakhin, Vladyslav
Veröffentlicht: (2026)
von: Parakhin, Vladyslav
Veröffentlicht: (2026)
Operational Memory Architecture for Kubernetes:Preserving Causal Context Across the Evidence Horizon
von: Khan, Shamsher
Veröffentlicht: (2026)
von: Khan, Shamsher
Veröffentlicht: (2026)
LayerKV: Optimizing Large Language Model Serving with Layer-wise KV Cache Management
von: Xiong, Yi, et al.
Veröffentlicht: (2024)
von: Xiong, Yi, et al.
Veröffentlicht: (2024)
A Framework for testing Federated Learning algorithms using an edge-like environment
von: Schwanck, Felipe Machado, et al.
Veröffentlicht: (2024)
von: Schwanck, Felipe Machado, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FLEdge: Benchmarking Federated Machine Learning Applications in Edge Computing Systems
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2023) -
How Can We Train Deep Learning Models Across Clouds and Continents? An Experimental Study
von: Erben, Alexander, et al.
Veröffentlicht: (2023) -
A Survey on Efficient Federated Learning Methods for Foundation Model Training
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2024) -
Scheduling the Unschedulable: Taming Black-Box LLM Inference at Scale
von: Yuan, Renzhong, et al.
Veröffentlicht: (2026) -
Designing Scalable Rate Limiting Systems: Algorithms, Architecture, and Distributed Solutions
von: Guan, Bo
Veröffentlicht: (2026)