LoDAdaC: a unified local training-based decentralized framework with adaptive gradients and compressed communication
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Wei, Panda, Anweshit, Pandey, Ujwal, Cook, Haven, Slota, George M., Wang, Naigang, Chen, Jie, Xu, Yangyang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Anonymized Network Sensing using C++26 std::execution on GPUs
por: Mandulak, Michael, et al.
Publicado: (2025)
por: Mandulak, Michael, et al.
Publicado: (2025)
HeLoCo: Efficient asynchronous low-communication training under data and device heterogeneity
por: Asif, Abdullah Al, et al.
Publicado: (2026)
por: Asif, Abdullah Al, et al.
Publicado: (2026)
Controlled disagreement improves generalization in decentralized training
por: Wang, Zesen, et al.
Publicado: (2026)
por: Wang, Zesen, et al.
Publicado: (2026)
From promise to practice: realizing high-performance decentralized training
por: Wang, Zesen, et al.
Publicado: (2024)
por: Wang, Zesen, et al.
Publicado: (2024)
A unified framework to improve the interoperability between HPC and Big Data languages and programming models
por: Piñeiro, César, et al.
Publicado: (2021)
por: Piñeiro, César, et al.
Publicado: (2021)
InfiniLoRA: Disaggregated Multi-LoRA Serving for Large Language Models
por: Chen, Hongyu, et al.
Publicado: (2026)
por: Chen, Hongyu, et al.
Publicado: (2026)
Dave: a decentralized, secure, and lively fraud-proof algorithm
por: Nehab, Diego, et al.
Publicado: (2024)
por: Nehab, Diego, et al.
Publicado: (2024)
PCCL: Photonic circuit-switched collective communication for distributed ML
por: Kumar, Abhishek Vijaya, et al.
Publicado: (2025)
por: Kumar, Abhishek Vijaya, et al.
Publicado: (2025)
Semi-decentralized Federated Time Series Prediction with Client Availability Budgets
por: Bao, Yunkai, et al.
Publicado: (2025)
por: Bao, Yunkai, et al.
Publicado: (2025)
CD-Raft: Reducing the Latency of Distributed Consensus in Cross-Domain Sites
por: Wang, Yangyang, et al.
Publicado: (2026)
por: Wang, Yangyang, et al.
Publicado: (2026)
Gathering of asynchronous robots on circle with limited visibility using finite communication
por: Sharma, Avisek, et al.
Publicado: (2025)
por: Sharma, Avisek, et al.
Publicado: (2025)
An Efficient Approach for Energy Conservation in Cloud Computing Environment
por: Pande, Sohan Kumar, et al.
Publicado: (2025)
por: Pande, Sohan Kumar, et al.
Publicado: (2025)
Self-adaptive, Requirements-driven Autoscaling of Microservices
por: Nunes, João Paulo Karol Santos, et al.
Publicado: (2024)
por: Nunes, João Paulo Karol Santos, et al.
Publicado: (2024)
Unlocking Real-Time Fluorescence Lifetime Imaging: Multi-Pixel Parallelism for FPGA-Accelerated Processing
por: Erbas, Ismail, et al.
Publicado: (2024)
por: Erbas, Ismail, et al.
Publicado: (2024)
Federated Learning framework for LoRaWAN-enabled IIoT communication: A case study
por: Sanchez, Oscar Torres, et al.
Publicado: (2024)
por: Sanchez, Oscar Torres, et al.
Publicado: (2024)
Decentralized and Self-adaptive Core Maintenance on Temporal Graphs
por: Rucci, Davide, et al.
Publicado: (2025)
por: Rucci, Davide, et al.
Publicado: (2025)
Application-level observability for adaptive Edge to Cloud continuum systems
por: Sidi, Kaddour, et al.
Publicado: (2026)
por: Sidi, Kaddour, et al.
Publicado: (2026)
Tasking framework for Adaptive Speculative Parallel Mesh Generation
por: Tsolakis, Christos, et al.
Publicado: (2024)
por: Tsolakis, Christos, et al.
Publicado: (2024)
A common parallel framework for LLP combinatorial problems
por: Alves, David Ribeiro, et al.
Publicado: (2026)
por: Alves, David Ribeiro, et al.
Publicado: (2026)
Scrutiny new framework in integrated distributed reliable systems
por: Gashti, Mehdi Zekriyapanah
Publicado: (2025)
por: Gashti, Mehdi Zekriyapanah
Publicado: (2025)
Characterizing Communication Patterns in Distributed Large Language Model Inference
por: Xu, Lang, et al.
Publicado: (2025)
por: Xu, Lang, et al.
Publicado: (2025)
Fog enabled distributed training architecture for federated learning
por: Kumar, Aditya, et al.
Publicado: (2024)
por: Kumar, Aditya, et al.
Publicado: (2024)
A Decentralized Microservice Scheduling Approach Using Service Mesh in Cloud-Edge Systems
por: Wen, Yangyang, et al.
Publicado: (2025)
por: Wen, Yangyang, et al.
Publicado: (2025)
Predictive-LoRA: A Proactive and Fragmentation-Aware Serverless Inference System for LLMs
por: Ni, Yinan, et al.
Publicado: (2025)
por: Ni, Yinan, et al.
Publicado: (2025)
EcoLoRA: Communication-Efficient Federated Fine-Tuning of Large Language Models
por: Liu, Han, et al.
Publicado: (2025)
por: Liu, Han, et al.
Publicado: (2025)
LoRA-C: Parameter-Efficient Fine-Tuning of Robust CNN for IoT Devices
por: Ding, Chuntao, et al.
Publicado: (2024)
por: Ding, Chuntao, et al.
Publicado: (2024)
emucxl: an emulation framework for CXL-based disaggregated memory applications
por: Gond, Raja, et al.
Publicado: (2024)
por: Gond, Raja, et al.
Publicado: (2024)
Simulating LLM training workloads for heterogeneous compute and network infrastructure
por: Kumar, Sumit, et al.
Publicado: (2025)
por: Kumar, Sumit, et al.
Publicado: (2025)
LRScheduler: A Layer-aware and Resource-adaptive Container Scheduler in Edge Computing
por: Tang, Zhiqing, et al.
Publicado: (2025)
por: Tang, Zhiqing, et al.
Publicado: (2025)
CaraServe: CPU-Assisted and Rank-Aware LoRA Serving for Generative LLM Inference
por: Li, Suyi, et al.
Publicado: (2024)
por: Li, Suyi, et al.
Publicado: (2024)
Revisiting Speculative Leaderless Protocols for Low-Latency BFT Replication
por: Qian, Daniel, et al.
Publicado: (2026)
por: Qian, Daniel, et al.
Publicado: (2026)
COPUS: Co-adaptive Parallelism and Batch Size Selection in Large Language Model Training
por: Sakip, Akhmed, et al.
Publicado: (2026)
por: Sakip, Akhmed, et al.
Publicado: (2026)
FedQuad: Adaptive Layer-wise LoRA Deployment and Activation Quantization for Federated Fine-Tuning
por: Li, Rukuo, et al.
Publicado: (2025)
por: Li, Rukuo, et al.
Publicado: (2025)
FDLoRA: Personalized Federated Learning of Large Language Model via Dual LoRA Tuning
por: QI, Jiaxing, et al.
Publicado: (2024)
por: QI, Jiaxing, et al.
Publicado: (2024)
GeoNimbus: A serverless framework to build earth observation and environmental services
por: Sánchez-Gallegos, Dante D., et al.
Publicado: (2025)
por: Sánchez-Gallegos, Dante D., et al.
Publicado: (2025)
FIRED: a fine-grained robust performance diagnosis framework for cloud applications
por: Xin, Ruyue, et al.
Publicado: (2022)
por: Xin, Ruyue, et al.
Publicado: (2022)
Flotilla: A scalable, modular and resilient federated learning framework for heterogeneous resources
por: Banerjee, Roopkatha, et al.
Publicado: (2025)
por: Banerjee, Roopkatha, et al.
Publicado: (2025)
Dflow, a Python framework for constructing cloud-native AI-for-Science workflows
por: Liu, Xinzijian, et al.
Publicado: (2024)
por: Liu, Xinzijian, et al.
Publicado: (2024)
LoHan: Low-Cost High-Performance Framework to Fine-Tune 100B Model on a Consumer GPU
por: Liao, Changyue, et al.
Publicado: (2024)
por: Liao, Changyue, et al.
Publicado: (2024)
An Overview on the Landscape of Self-Adaptive Cloud Design and Operation Patterns: Goals, Strategies, Tooling, Evaluation, and Dataset Perspectives
por: Angelis, Apostolos, et al.
Publicado: (2025)
por: Angelis, Apostolos, et al.
Publicado: (2025)
Ejemplares similares
-
Anonymized Network Sensing using C++26 std::execution on GPUs
por: Mandulak, Michael, et al.
Publicado: (2025) -
HeLoCo: Efficient asynchronous low-communication training under data and device heterogeneity
por: Asif, Abdullah Al, et al.
Publicado: (2026) -
Controlled disagreement improves generalization in decentralized training
por: Wang, Zesen, et al.
Publicado: (2026) -
From promise to practice: realizing high-performance decentralized training
por: Wang, Zesen, et al.
Publicado: (2024) -
A unified framework to improve the interoperability between HPC and Big Data languages and programming models
por: Piñeiro, César, et al.
Publicado: (2021)