Towards Decentralized and Sustainable Foundation Model Training with the Edge
Fuente:
arXiv
Saved in:
| Main Authors: | Xue, Leyang, Madhyastha, Meghana, Burns, Randal, Lee, Myungjin, Marina, Mahesh K. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Harnessing Idle Compute at the Edge for Foundation Model Training
by: Xue, Leyang, et al.
Published: (2025)
by: Xue, Leyang, et al.
Published: (2025)
Masked Matrix Multiplication for Emergent Sparsity
by: Wheatman, Brian, et al.
Published: (2024)
by: Wheatman, Brian, et al.
Published: (2024)
HybridServe: Efficient Serving of Large AI Models with Confidence-Based Cascade Routing
by: Xue, Leyang, et al.
Published: (2025)
by: Xue, Leyang, et al.
Published: (2025)
TUBO: A Tailored ML Framework for Reliable Network Traffic Forecasting
by: Yuan, Zhihang, et al.
Published: (2026)
by: Yuan, Zhihang, et al.
Published: (2026)
MoE-Infinity: Efficient MoE Inference on Personal Machines with Sparsity-Aware Expert Cache
by: Xue, Leyang, et al.
Published: (2024)
by: Xue, Leyang, et al.
Published: (2024)
To 2:4 Sparsity and Beyond: Neuron-level Activation Function to Accelerate LLM Pre-Training
by: Madhyastha, Meghana, et al.
Published: (2026)
by: Madhyastha, Meghana, et al.
Published: (2026)
Edge-Parallel Graph Encoder Embedding
by: Lubonja, Ariel, et al.
Published: (2024)
by: Lubonja, Ariel, et al.
Published: (2024)
A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility
by: Madhyastha, Pranava
Published: (2026)
by: Madhyastha, Pranava
Published: (2026)
A Collaborative Process Parameter Recommender System for Fleets of Networked Manufacturing Machines -- with Application to 3D Printing
by: Wang, Weishi, et al.
Published: (2025)
by: Wang, Weishi, et al.
Published: (2025)
Curvature Tuning: Provable Training-free Model Steering From a Single Parameter
by: Hu, Leyang, et al.
Published: (2025)
by: Hu, Leyang, et al.
Published: (2025)
SeedFlood: A Step Toward Scalable Decentralized Training of LLMs
by: Kim, Jihun, et al.
Published: (2026)
by: Kim, Jihun, et al.
Published: (2026)
LIFL: A Lightweight, Event-driven Serverless Platform for Federated Learning
by: Qi, Shixiong, et al.
Published: (2024)
by: Qi, Shixiong, et al.
Published: (2024)
Linear Independence of Generalized Neurons and Related Functions
by: Zhang, Leyang
Published: (2024)
by: Zhang, Leyang
Published: (2024)
Working Memory Constraints Scaffold Learning in Transformers under Data Scarcity
by: Madhyastha, Pranava, et al.
Published: (2026)
by: Madhyastha, Pranava, et al.
Published: (2026)
FedAuxHMTL: Federated Auxiliary Hard-Parameter Sharing Multi-Task Learning for Network Edge Traffic Classification
by: Ahmed, Faisal, et al.
Published: (2024)
by: Ahmed, Faisal, et al.
Published: (2024)
Guiding Time-Varying Generative Models with Natural Gradients on Exponential Family Manifold
by: Liu, Song, et al.
Published: (2025)
by: Liu, Song, et al.
Published: (2025)
Improved identification of breakpoints in piecewise regression and its applications
by: Kim, Taehyeong, et al.
Published: (2024)
by: Kim, Taehyeong, et al.
Published: (2024)
$\texttt{SEM-CTRL}$: Semantically Controlled Decoding
by: Albinhassan, Mohammad, et al.
Published: (2025)
by: Albinhassan, Mohammad, et al.
Published: (2025)
Composer: A Search Framework for Hybrid Neural Architecture Design
by: Acun, Bilge, et al.
Published: (2025)
by: Acun, Bilge, et al.
Published: (2025)
DεpS: Delayed ε-Shrinking for Faster Once-For-All Training
by: Annavajjala, Aditya, et al.
Published: (2024)
by: Annavajjala, Aditya, et al.
Published: (2024)
Making MoE-based LLM Inference Resilient with Tarragon
by: Zhang, Songyu, et al.
Published: (2026)
by: Zhang, Songyu, et al.
Published: (2026)
Towards Large-Scale Training of Pathology Foundation Models
by: ai, kaiko., et al.
Published: (2024)
by: ai, kaiko., et al.
Published: (2024)
Protocol Models: Scaling Decentralized Training with Communication-Efficient Model Parallelism
by: Ramasinghe, Sameera, et al.
Published: (2025)
by: Ramasinghe, Sameera, et al.
Published: (2025)
THEMIS: Unlocking Pretrained Knowledge with Foundation Model Embeddings for Anomaly Detection in Time Series
by: Lorik, Yadav Mahesh, et al.
Published: (2025)
by: Lorik, Yadav Mahesh, et al.
Published: (2025)
DeCAF: Decentralized Consensus-And-Factorization for Low-Rank Adaptation of Foundation Models
by: Saadati, Nastaran, et al.
Published: (2025)
by: Saadati, Nastaran, et al.
Published: (2025)
AltTS: A Dual-Path Framework with Alternating Optimization for Multivariate Time Series Forecasting
by: Yuan, Zhihang, et al.
Published: (2026)
by: Yuan, Zhihang, et al.
Published: (2026)
Towards Heterogeneity-Aware and Energy-Efficient Topology Optimization for Decentralized Federated Learning in Edge Environment
by: Liu, Yuze, et al.
Published: (2025)
by: Liu, Yuze, et al.
Published: (2025)
QuantFL: Sustainable Federated Learning for Edge IoT via Pre-Trained Model Quantisation
by: Herath, Charuka, et al.
Published: (2026)
by: Herath, Charuka, et al.
Published: (2026)
How Foundational are Foundation Models for Time Series Forecasting?
by: Karaouli, Nouha, et al.
Published: (2025)
by: Karaouli, Nouha, et al.
Published: (2025)
Personalized Federated Fine-tuning for Heterogeneous Data: An Automatic Rank Learning Approach via Two-Level LoRA
by: Hao, Jie, et al.
Published: (2025)
by: Hao, Jie, et al.
Published: (2025)
Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models
by: Hübotter, Jonas, et al.
Published: (2025)
by: Hübotter, Jonas, et al.
Published: (2025)
Toward a Graph Foundation Model: Pre-Training Transformers With Random Walks
by: Tang, Ziyuan, et al.
Published: (2025)
by: Tang, Ziyuan, et al.
Published: (2025)
Uncovering Critical Sets of Deep Neural Networks via Sample-Independent Critical Lifting
by: Zhang, Leyang, et al.
Published: (2025)
by: Zhang, Leyang, et al.
Published: (2025)
Prediction-Assisted Online Distributed Deep Learning Workload Scheduling in GPU Clusters
by: Luo, Ziyue, et al.
Published: (2025)
by: Luo, Ziyue, et al.
Published: (2025)
Towards Graph Foundation Models: Training on Knowledge Graphs Enables Transferability to General Graphs
by: Wang, Kai, et al.
Published: (2024)
by: Wang, Kai, et al.
Published: (2024)
Toward Understanding BERT-Like Pre-Training for DNA Foundation Models
by: Liang, Chaoqi, et al.
Published: (2023)
by: Liang, Chaoqi, et al.
Published: (2023)
EdgeServe: A Streaming System for Decentralized Model Serving
by: Shaowang, Ted, et al.
Published: (2023)
by: Shaowang, Ted, et al.
Published: (2023)
Learning and Enforcing Context-Sensitive Control for LLMs
by: Albinhassan, Mohammad, et al.
Published: (2026)
by: Albinhassan, Mohammad, et al.
Published: (2026)
Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems
by: Karnam, Meghana, et al.
Published: (2026)
by: Karnam, Meghana, et al.
Published: (2026)
Critical Patch-Aware Sparse Prompting with Decoupled Training for Continual Learning on the Edge
by: Lim, Wonseon, et al.
Published: (2026)
by: Lim, Wonseon, et al.
Published: (2026)
Similar Items
-
On Harnessing Idle Compute at the Edge for Foundation Model Training
by: Xue, Leyang, et al.
Published: (2025) -
Masked Matrix Multiplication for Emergent Sparsity
by: Wheatman, Brian, et al.
Published: (2024) -
HybridServe: Efficient Serving of Large AI Models with Confidence-Based Cascade Routing
by: Xue, Leyang, et al.
Published: (2025) -
TUBO: A Tailored ML Framework for Reliable Network Traffic Forecasting
by: Yuan, Zhihang, et al.
Published: (2026) -
MoE-Infinity: Efficient MoE Inference on Personal Machines with Sparsity-Aware Expert Cache
by: Xue, Leyang, et al.
Published: (2024)