Salvato in:
| Autori principali: | Tang, Weiheng, Li, Jingyi, Chen, Lin, Chen, Xu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.10831 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Jupiter: Fast and Resource-Efficient Collaborative Inference of Generative LLMs on Edge Devices
di: Ye, Shengyuan, et al.
Pubblicazione: (2025)
di: Ye, Shengyuan, et al.
Pubblicazione: (2025)
HALO: Semantic-Aware Distributed LLM Inference in Lossy Edge Network
di: Zheng, Peirong, et al.
Pubblicazione: (2026)
di: Zheng, Peirong, et al.
Pubblicazione: (2026)
Adaptive Parameter-Efficient Federated Fine-Tuning on Heterogeneous Devices
di: Liu, Jun, et al.
Pubblicazione: (2024)
di: Liu, Jun, et al.
Pubblicazione: (2024)
Hierarchical Split Federated Learning: Convergence Analysis and System Optimization
di: Lin, Zheng, et al.
Pubblicazione: (2024)
di: Lin, Zheng, et al.
Pubblicazione: (2024)
Trust-Aware Routing for Distributed Generative AI Inference at the Edge
di: Nguyen, Chanh, et al.
Pubblicazione: (2026)
di: Nguyen, Chanh, et al.
Pubblicazione: (2026)
Towards Edge General Intelligence via Large Language Models: Opportunities and Challenges
di: Chen, Handi, et al.
Pubblicazione: (2024)
di: Chen, Handi, et al.
Pubblicazione: (2024)
Resource-Efficient Personal Large Language Models Fine-Tuning with Collaborative Edge Computing
di: Ye, Shengyuan, et al.
Pubblicazione: (2024)
di: Ye, Shengyuan, et al.
Pubblicazione: (2024)
Rina: Enhancing Ring-AllReduce with In-network Aggregation in Distributed Model Training
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
eACGM: Non-instrumented Performance Tracing and Anomaly Detection towards Machine Learning Systems
di: Xu, Ruilin, et al.
Pubblicazione: (2025)
di: Xu, Ruilin, et al.
Pubblicazione: (2025)
Edge Graph Intelligence: Reciprocally Empowering Edge Networks with Graph Intelligence
di: Zeng, Liekang, et al.
Pubblicazione: (2024)
di: Zeng, Liekang, et al.
Pubblicazione: (2024)
Optimizing Resource Allocation for Geographically-Distributed Inference by Large Language Models
di: Sun, Tingyang, et al.
Pubblicazione: (2025)
di: Sun, Tingyang, et al.
Pubblicazione: (2025)
Agentic Performance at the Edge: Insights from Benchmarking
di: Wang, Shiqiang, et al.
Pubblicazione: (2026)
di: Wang, Shiqiang, et al.
Pubblicazione: (2026)
Smaller, Smarter, Closer: The Edge of Collaborative Generative AI
di: Morabito, Roberto, et al.
Pubblicazione: (2025)
di: Morabito, Roberto, et al.
Pubblicazione: (2025)
Teola: Towards End-to-End Optimization of LLM-based Applications
di: Tan, Xin, et al.
Pubblicazione: (2024)
di: Tan, Xin, et al.
Pubblicazione: (2024)
Optimizing Split Learning Latency in TinyML-Based IoT Systems
di: Jenhani, Zied, et al.
Pubblicazione: (2025)
di: Jenhani, Zied, et al.
Pubblicazione: (2025)
Enabling Intelligent Vehicular Networks Through Distributed Learning in the Non-Terrestrial Networks 6G Vision
di: Naseh, David, et al.
Pubblicazione: (2023)
di: Naseh, David, et al.
Pubblicazione: (2023)
Communication Optimization for Decentralized Learning atop Bandwidth-limited Edge Networks
di: Sun, Tingyang, et al.
Pubblicazione: (2025)
di: Sun, Tingyang, et al.
Pubblicazione: (2025)
Self-Healing Network of Interconnected Edge Devices Empowered by Infrastructure-as-Code and LoRa Communication
di: Carson, Rob, et al.
Pubblicazione: (2025)
di: Carson, Rob, et al.
Pubblicazione: (2025)
Implementation of Big AI Models for Wireless Networks with Collaborative Edge Computing
di: Zeng, Liekang, et al.
Pubblicazione: (2024)
di: Zeng, Liekang, et al.
Pubblicazione: (2024)
ScaleAcross Explorer: Exploring Communication Optimization for Scale-Across AI Model Training
di: Li, Minghao, et al.
Pubblicazione: (2026)
di: Li, Minghao, et al.
Pubblicazione: (2026)
Digital Twinning of a Pressurized Water Reactor Startup Operation and Partial Computational Offloading in In-network Computing-Assisted Multiaccess Edge Computing
di: Aliyu, Ibrahim, et al.
Pubblicazione: (2024)
di: Aliyu, Ibrahim, et al.
Pubblicazione: (2024)
Quality-of-Service Aware LLM Routing for Edge Computing with Multiple Experts
di: Yang, Jin, et al.
Pubblicazione: (2025)
di: Yang, Jin, et al.
Pubblicazione: (2025)
Early-Exit meets Model-Distributed Inference at Edge Networks
di: Colocrese, Marco, et al.
Pubblicazione: (2024)
di: Colocrese, Marco, et al.
Pubblicazione: (2024)
SpaceMoE: Realizing Distributed Mixture-of-Experts Inference over Space Networks
di: Wang, Zhanwei, et al.
Pubblicazione: (2026)
di: Wang, Zhanwei, et al.
Pubblicazione: (2026)
Collective Communication Profiling of Modern-day Machine Learning Workloads
di: Gupta, Jit, et al.
Pubblicazione: (2025)
di: Gupta, Jit, et al.
Pubblicazione: (2025)
Generative AI on the Edge: Architecture and Performance Evaluation
di: Nezami, Zeinab, et al.
Pubblicazione: (2024)
di: Nezami, Zeinab, et al.
Pubblicazione: (2024)
Galaxy: A Resource-Efficient Collaborative Edge AI System for In-situ Transformer Inference
di: Ye, Shengyuan, et al.
Pubblicazione: (2024)
di: Ye, Shengyuan, et al.
Pubblicazione: (2024)
Edge-First Language Model Inference: Models, Metrics, and Tradeoffs
di: Jang, SiYoung, et al.
Pubblicazione: (2025)
di: Jang, SiYoung, et al.
Pubblicazione: (2025)
Topology-aware Microservice Architecture in Edge Networks: Deployment Optimization and Implementation
di: Chen, Yuang, et al.
Pubblicazione: (2025)
di: Chen, Yuang, et al.
Pubblicazione: (2025)
Context-Aware Orchestration of Energy-Efficient Gossip Learning Schemes
di: Dinani, Mina Aghaei, et al.
Pubblicazione: (2024)
di: Dinani, Mina Aghaei, et al.
Pubblicazione: (2024)
NebulaFL: Effective Asynchronous Federated Learning for JointCloud Computing
di: Gao, Fei, et al.
Pubblicazione: (2024)
di: Gao, Fei, et al.
Pubblicazione: (2024)
Enabling Reconfiguration-Communication Overlap for Collective Communication in Optical Networks
di: Wu, Changbo, et al.
Pubblicazione: (2025)
di: Wu, Changbo, et al.
Pubblicazione: (2025)
The Implications of Decentralization in Blockchained Federated Learning: Evaluating the Impact of Model Staleness and Inconsistencies
di: Wilhelmi, Francesc, et al.
Pubblicazione: (2023)
di: Wilhelmi, Francesc, et al.
Pubblicazione: (2023)
Federated Continual Learning for Edge-AI: A Comprehensive Survey
di: Wang, Zi, et al.
Pubblicazione: (2024)
di: Wang, Zi, et al.
Pubblicazione: (2024)
Resilient by Design -- Active Inference for Distributed Continuum Intelligence
di: Donta, Praveen Kumar, et al.
Pubblicazione: (2025)
di: Donta, Praveen Kumar, et al.
Pubblicazione: (2025)
Device Sampling and Resource Optimization for Federated Learning in Cooperative Edge Networks
di: Wang, Su, et al.
Pubblicazione: (2023)
di: Wang, Su, et al.
Pubblicazione: (2023)
A VM-HDL Co-Simulation Framework for Systems with PCIe-Connected FPGAs
di: Cho, Shenghsun, et al.
Pubblicazione: (2025)
di: Cho, Shenghsun, et al.
Pubblicazione: (2025)
Towards Net-Zero Carbon Emissions in Network AI for 6G and Beyond
di: Zhang, Peng, et al.
Pubblicazione: (2023)
di: Zhang, Peng, et al.
Pubblicazione: (2023)
Network Anomaly Detection in Distributed Edge Computing Infrastructure
di: Marfo, William, et al.
Pubblicazione: (2025)
di: Marfo, William, et al.
Pubblicazione: (2025)
Diffusion Models on the Edge: Challenges, Optimizations, and Applications
di: Zheng, Dongqi
Pubblicazione: (2025)
di: Zheng, Dongqi
Pubblicazione: (2025)
Documenti analoghi
-
Jupiter: Fast and Resource-Efficient Collaborative Inference of Generative LLMs on Edge Devices
di: Ye, Shengyuan, et al.
Pubblicazione: (2025) -
HALO: Semantic-Aware Distributed LLM Inference in Lossy Edge Network
di: Zheng, Peirong, et al.
Pubblicazione: (2026) -
Adaptive Parameter-Efficient Federated Fine-Tuning on Heterogeneous Devices
di: Liu, Jun, et al.
Pubblicazione: (2024) -
Hierarchical Split Federated Learning: Convergence Analysis and System Optimization
di: Lin, Zheng, et al.
Pubblicazione: (2024) -
Trust-Aware Routing for Distributed Generative AI Inference at the Edge
di: Nguyen, Chanh, et al.
Pubblicazione: (2026)