Enregistré dans:
| Auteurs principaux: | Mishra, Asit, Stosic, Dusan, Layton, Simon, Micikevicius, Paulius |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2506.08027 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Future of Large Language Model Pre-training is Federated
par: Sani, Lorenzo, et autres
Publié: (2024)
par: Sani, Lorenzo, et autres
Publié: (2024)
Improving training time and GPU utilization in geo-distributed language model training
par: Palak, et autres
Publié: (2024)
par: Palak, et autres
Publié: (2024)
Data movement limits to frontier model training
par: Erdil, Ege, et autres
Publié: (2024)
par: Erdil, Ege, et autres
Publié: (2024)
Pre-Deployment Complexity Estimation for Federated Perception Systems
par: Solaiman, KMA, et autres
Publié: (2026)
par: Solaiman, KMA, et autres
Publié: (2026)
FLStore: Efficient Federated Learning Storage for non-training workloads
par: Khan, Ahmad Faraz, et autres
Publié: (2025)
par: Khan, Ahmad Faraz, et autres
Publié: (2025)
LlamaDuo: LLMOps Pipeline for Seamless Migration from Service LLMs to Small-Scale Local LLMs
par: Park, Chansung, et autres
Publié: (2024)
par: Park, Chansung, et autres
Publié: (2024)
Marconi: Prefix Caching for the Era of Hybrid LLMs
par: Pan, Rui, et autres
Publié: (2024)
par: Pan, Rui, et autres
Publié: (2024)
MinT: Managed Infrastructure for Training and Serving Millions of LLMs
par: Lab, Mind, et autres
Publié: (2026)
par: Lab, Mind, et autres
Publié: (2026)
LoRAFusion: Efficient LoRA Fine-Tuning for LLMs
par: Zhu, Zhanda, et autres
Publié: (2025)
par: Zhu, Zhanda, et autres
Publié: (2025)
SFPrompt: Communication-Efficient Split Federated Fine-Tuning for Large Pre-Trained Models over Resource-Limited Devices
par: Cao, Linxiao, et autres
Publié: (2024)
par: Cao, Linxiao, et autres
Publié: (2024)
DISTFLASHATTN: Distributed Memory-efficient Attention for Long-context LLMs Training
par: Li, Dacheng, et autres
Publié: (2023)
par: Li, Dacheng, et autres
Publié: (2023)
Fail Fast, Win Big: Rethinking the Drafting Strategy in Speculative Decoding via Diffusion LLMs
par: Pan, Rui, et autres
Publié: (2025)
par: Pan, Rui, et autres
Publié: (2025)
Towards the Next Frontier of LLMs, Training on Private Data: A Cross-Domain Benchmark for Federated Fine-Tuning
par: Jimenez-Gutierrez, Daniel M., et autres
Publié: (2026)
par: Jimenez-Gutierrez, Daniel M., et autres
Publié: (2026)
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving
par: Hu, Huanqi, et autres
Publié: (2025)
par: Hu, Huanqi, et autres
Publié: (2025)
TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training
par: Liang, Wanchao, et autres
Publié: (2024)
par: Liang, Wanchao, et autres
Publié: (2024)
Bitnet.cpp: Efficient Edge Inference for Ternary LLMs
par: Wang, Jinheng, et autres
Publié: (2025)
par: Wang, Jinheng, et autres
Publié: (2025)
Vertical Federated Learning Hybrid Local Pre-training
par: Li, Wenguo, et autres
Publié: (2024)
par: Li, Wenguo, et autres
Publié: (2024)
Poison Once, Refuse Forever: Weaponizing Alignment for Injecting Bias in LLMs
par: Mamun, Md Abdullah Al, et autres
Publié: (2025)
par: Mamun, Md Abdullah Al, et autres
Publié: (2025)
Low-Rank GEMM: Efficient Matrix Multiplication via Low-Rank Approximation with FP8 Acceleration
par: Metere, Alfredo
Publié: (2025)
par: Metere, Alfredo
Publié: (2025)
A Structure-Agnostic Co-Tuning Framework for LLMs and SLMs in Cloud-Edge Systems
par: Liu, Yuze, et autres
Publié: (2025)
par: Liu, Yuze, et autres
Publié: (2025)
SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs
par: Cao, Yadi, et autres
Publié: (2026)
par: Cao, Yadi, et autres
Publié: (2026)
GPT-FL: Generative Pre-trained Model-Assisted Federated Learning
par: Zhang, Tuo, et autres
Publié: (2023)
par: Zhang, Tuo, et autres
Publié: (2023)
A Unified Convergence Analysis for Semi-Decentralized Learning: Sampled-to-Sampled vs. Sampled-to-All Communication
par: Rodio, Angelo, et autres
Publié: (2025)
par: Rodio, Angelo, et autres
Publié: (2025)
Synera: Synergistic LLM Serving across Device and Cloud at Scale
par: Wang, Genglin, et autres
Publié: (2025)
par: Wang, Genglin, et autres
Publié: (2025)
PGT-I: Scaling Spatiotemporal GNNs with Memory-Efficient Distributed Training
par: Ockerman, Seth, et autres
Publié: (2025)
par: Ockerman, Seth, et autres
Publié: (2025)
A-3PO: Accelerating Asynchronous LLM Training with Staleness-aware Proximal Policy Approximation
par: Li, Xiaocan, et autres
Publié: (2025)
par: Li, Xiaocan, et autres
Publié: (2025)
Topology-Aware Knowledge Propagation in Decentralized Learning
par: Sakarvadia, Mansi, et autres
Publié: (2025)
par: Sakarvadia, Mansi, et autres
Publié: (2025)
FedHERO: A Federated Learning Approach for Node Classification Task on Heterophilic Graphs
par: Chen, Zihan, et autres
Publié: (2025)
par: Chen, Zihan, et autres
Publié: (2025)
FedCGD: Collective Gradient Divergence Optimized Scheduling for Wireless Federated Learning
par: Chen, Tan, et autres
Publié: (2025)
par: Chen, Tan, et autres
Publié: (2025)
Research on Edge Computing and Cloud Collaborative Resource Scheduling Optimization Based on Deep Reinforcement Learning
par: Wang, Yuqing, et autres
Publié: (2025)
par: Wang, Yuqing, et autres
Publié: (2025)
HSplitLoRA: A Heterogeneous Split Parameter-Efficient Fine-Tuning Framework for Large Language Models
par: Lin, Zheng, et autres
Publié: (2025)
par: Lin, Zheng, et autres
Publié: (2025)
Learning Like Humans: Resource-Efficient Federated Fine-Tuning through Cognitive Developmental Stages
par: Wu, Yebo, et autres
Publié: (2025)
par: Wu, Yebo, et autres
Publié: (2025)
PipeOffload: Improving Scalability of Pipeline Parallelism with Memory Optimization
par: Wan, Xinyi, et autres
Publié: (2025)
par: Wan, Xinyi, et autres
Publié: (2025)
Accelerating Privacy-Preserving Federated Learning in Large-Scale LEO Satellite Systems
par: Guo, Binquan, et autres
Publié: (2025)
par: Guo, Binquan, et autres
Publié: (2025)
On Using Large-Batches in Federated Learning
par: Tyagi, Sahil
Publié: (2025)
par: Tyagi, Sahil
Publié: (2025)
Adaptive Approach to Enhance Machine Learning Scheduling Algorithms During Runtime Using Reinforcement Learning in Metascheduling Applications
par: Alshaer, Samer, et autres
Publié: (2025)
par: Alshaer, Samer, et autres
Publié: (2025)
Federated Attention: A Distributed Paradigm for Collaborative LLM Inference over Edge Networks
par: Deng, Xiumei, et autres
Publié: (2025)
par: Deng, Xiumei, et autres
Publié: (2025)
Adaptive Graph Pruning with Sudden-Events Evaluation for Traffic Prediction using Online Semi-Decentralized ST-GNNs
par: Kralj, Ivan, et autres
Publié: (2025)
par: Kralj, Ivan, et autres
Publié: (2025)
Task-Agnostic Federation over Decentralized Data: Research Landscape and Visions
par: Wu, Wentai, et autres
Publié: (2025)
par: Wu, Wentai, et autres
Publié: (2025)
DPQuant: Efficient and Differentially-Private Model Training via Dynamic Quantization Scheduling
par: Gao, Yubo, et autres
Publié: (2025)
par: Gao, Yubo, et autres
Publié: (2025)
Documents similaires
-
The Future of Large Language Model Pre-training is Federated
par: Sani, Lorenzo, et autres
Publié: (2024) -
Improving training time and GPU utilization in geo-distributed language model training
par: Palak, et autres
Publié: (2024) -
Data movement limits to frontier model training
par: Erdil, Ege, et autres
Publié: (2024) -
Pre-Deployment Complexity Estimation for Federated Perception Systems
par: Solaiman, KMA, et autres
Publié: (2026) -
FLStore: Efficient Federated Learning Storage for non-training workloads
par: Khan, Ahmad Faraz, et autres
Publié: (2025)