Adaptive Rank Allocation for Federated Parameter-Efficient Fine-Tuning of Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Fei, Hu, Jia, Min, Geyong, Wang, Shiqiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Parameter-Efficient Federated Fine-Tuning on Heterogeneous Devices
by: Liu, Jun, et al.
Published: (2024)
by: Liu, Jun, et al.
Published: (2024)
Federated Continual Learning for Edge-AI: A Comprehensive Survey
by: Wang, Zi, et al.
Published: (2024)
by: Wang, Zi, et al.
Published: (2024)
Agentic Performance at the Edge: Insights from Benchmarking
by: Wang, Shiqiang, et al.
Published: (2026)
by: Wang, Shiqiang, et al.
Published: (2026)
Optimizing Resource Allocation for Geographically-Distributed Inference by Large Language Models
by: Sun, Tingyang, et al.
Published: (2025)
by: Sun, Tingyang, et al.
Published: (2025)
Resource-Efficient Personal Large Language Models Fine-Tuning with Collaborative Edge Computing
by: Ye, Shengyuan, et al.
Published: (2024)
by: Ye, Shengyuan, et al.
Published: (2024)
Preventing Rank Collapse in Federated Low-Rank Adaptation with Client Heterogeneity
by: Wu, Fei, et al.
Published: (2026)
by: Wu, Fei, et al.
Published: (2026)
NebulaFL: Effective Asynchronous Federated Learning for JointCloud Computing
by: Gao, Fei, et al.
Published: (2024)
by: Gao, Fei, et al.
Published: (2024)
The Implications of Decentralization in Blockchained Federated Learning: Evaluating the Impact of Model Staleness and Inconsistencies
by: Wilhelmi, Francesc, et al.
Published: (2023)
by: Wilhelmi, Francesc, et al.
Published: (2023)
Towards Edge General Intelligence via Large Language Models: Opportunities and Challenges
by: Chen, Handi, et al.
Published: (2024)
by: Chen, Handi, et al.
Published: (2024)
XWind: A Cross-site Router for Large Language Model Inference Serving at Renewable Energy Farms
by: Reddy, Tella Rajashekhar, et al.
Published: (2026)
by: Reddy, Tella Rajashekhar, et al.
Published: (2026)
Context-Aware Orchestration of Energy-Efficient Gossip Learning Schemes
by: Dinani, Mina Aghaei, et al.
Published: (2024)
by: Dinani, Mina Aghaei, et al.
Published: (2024)
Rina: Enhancing Ring-AllReduce with In-network Aggregation in Distributed Model Training
by: Chen, Zixuan, et al.
Published: (2024)
by: Chen, Zixuan, et al.
Published: (2024)
Jupiter: Fast and Resource-Efficient Collaborative Inference of Generative LLMs on Edge Devices
by: Ye, Shengyuan, et al.
Published: (2025)
by: Ye, Shengyuan, et al.
Published: (2025)
KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving
by: Liu, Zedong, et al.
Published: (2026)
by: Liu, Zedong, et al.
Published: (2026)
SpaceMoE: Realizing Distributed Mixture-of-Experts Inference over Space Networks
by: Wang, Zhanwei, et al.
Published: (2026)
by: Wang, Zhanwei, et al.
Published: (2026)
ScaleAcross Explorer: Exploring Communication Optimization for Scale-Across AI Model Training
by: Li, Minghao, et al.
Published: (2026)
by: Li, Minghao, et al.
Published: (2026)
Resource Allocation Driven by Large Models in Future Semantic-Aware Networks
by: Zhang, Haijun, et al.
Published: (2025)
by: Zhang, Haijun, et al.
Published: (2025)
Enabling Reconfiguration-Communication Overlap for Collective Communication in Optical Networks
by: Wu, Changbo, et al.
Published: (2025)
by: Wu, Changbo, et al.
Published: (2025)
Edge-First Language Model Inference: Models, Metrics, and Tradeoffs
by: Jang, SiYoung, et al.
Published: (2025)
by: Jang, SiYoung, et al.
Published: (2025)
A Generative Caching System for Large Language Models
by: Iyengar, Arun, et al.
Published: (2025)
by: Iyengar, Arun, et al.
Published: (2025)
Satellite Federated Fine-Tuning for Foundation Models in Space Computing Power Networks
by: Zhu, Yan, et al.
Published: (2025)
by: Zhu, Yan, et al.
Published: (2025)
ProFe: Communication-Efficient Decentralized Federated Learning via Distillation and Prototypes
by: Sánchez, Pedro Miguel Sánchez, et al.
Published: (2024)
by: Sánchez, Pedro Miguel Sánchez, et al.
Published: (2024)
Federated Learning and Evolutionary Game Model for Fog Federation Formation
by: Yasser, Zyad, et al.
Published: (2024)
by: Yasser, Zyad, et al.
Published: (2024)
Collective Communication Profiling of Modern-day Machine Learning Workloads
by: Gupta, Jit, et al.
Published: (2025)
by: Gupta, Jit, et al.
Published: (2025)
Digital Twinning of a Pressurized Water Reactor Startup Operation and Partial Computational Offloading in In-network Computing-Assisted Multiaccess Edge Computing
by: Aliyu, Ibrahim, et al.
Published: (2024)
by: Aliyu, Ibrahim, et al.
Published: (2024)
HALO: Semantic-Aware Distributed LLM Inference in Lossy Edge Network
by: Zheng, Peirong, et al.
Published: (2026)
by: Zheng, Peirong, et al.
Published: (2026)
Towards Net-Zero Carbon Emissions in Network AI for 6G and Beyond
by: Zhang, Peng, et al.
Published: (2023)
by: Zhang, Peng, et al.
Published: (2023)
Optimizing Split Learning Latency in TinyML-Based IoT Systems
by: Jenhani, Zied, et al.
Published: (2025)
by: Jenhani, Zied, et al.
Published: (2025)
Trust-Aware Routing for Distributed Generative AI Inference at the Edge
by: Nguyen, Chanh, et al.
Published: (2026)
by: Nguyen, Chanh, et al.
Published: (2026)
When IoT Meet LLMs: Applications and Challenges
by: Kok, Ibrahim, et al.
Published: (2024)
by: Kok, Ibrahim, et al.
Published: (2024)
When Digital Twin Meets 6G: Concepts, Obstacles, and Research Prospects
by: Liu, Wenshuai, et al.
Published: (2024)
by: Liu, Wenshuai, et al.
Published: (2024)
Design and Optimization of Hierarchical Gradient Coding for Distributed Learning at Edge Devices
by: Tang, Weiheng, et al.
Published: (2024)
by: Tang, Weiheng, et al.
Published: (2024)
AI Greenferencing: Routing AI Inferencing to Green Modular Data Centers with Heron
by: Reddy, Tella Rajashekhar, et al.
Published: (2025)
by: Reddy, Tella Rajashekhar, et al.
Published: (2025)
Cluster Topology-Driven Placement of Experts Reduces Network Traffic in MoE Inference
by: Sivtsov, Danil, et al.
Published: (2025)
by: Sivtsov, Danil, et al.
Published: (2025)
High-speed Networking for Giga-Scale AI Factories
by: Khashab, Sajy, et al.
Published: (2026)
by: Khashab, Sajy, et al.
Published: (2026)
Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics
by: Ma, Bole, et al.
Published: (2026)
by: Ma, Bole, et al.
Published: (2026)
Smaller, Smarter, Closer: The Edge of Collaborative Generative AI
by: Morabito, Roberto, et al.
Published: (2025)
by: Morabito, Roberto, et al.
Published: (2025)
Teola: Towards End-to-End Optimization of LLM-based Applications
by: Tan, Xin, et al.
Published: (2024)
by: Tan, Xin, et al.
Published: (2024)
eACGM: Non-instrumented Performance Tracing and Anomaly Detection towards Machine Learning Systems
by: Xu, Ruilin, et al.
Published: (2025)
by: Xu, Ruilin, et al.
Published: (2025)
An Open API Architecture to Discover the Trustworthy Explanation of Cloud AI Services
by: Wang, Zerui, et al.
Published: (2024)
by: Wang, Zerui, et al.
Published: (2024)
Similar Items
-
Adaptive Parameter-Efficient Federated Fine-Tuning on Heterogeneous Devices
by: Liu, Jun, et al.
Published: (2024) -
Federated Continual Learning for Edge-AI: A Comprehensive Survey
by: Wang, Zi, et al.
Published: (2024) -
Agentic Performance at the Edge: Insights from Benchmarking
by: Wang, Shiqiang, et al.
Published: (2026) -
Optimizing Resource Allocation for Geographically-Distributed Inference by Large Language Models
by: Sun, Tingyang, et al.
Published: (2025) -
Resource-Efficient Personal Large Language Models Fine-Tuning with Collaborative Edge Computing
by: Ye, Shengyuan, et al.
Published: (2024)