Trivance: Latency-Optimal AllReduce by Shortcutting Multiport Networks
Fuente:
arXiv
Salvato in:
| Autori principali: | Juerss, Anton, Addanki, Vamsi, Schmid, Stefan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Short-circuiting Rings for Low-Latency AllReduce
di: Hammer, Sarah-Michelle, et al.
Pubblicazione: (2025)
di: Hammer, Sarah-Michelle, et al.
Pubblicazione: (2025)
Revisiting Bruck: Phase-Efficient All-to-All Communication in Reconfigurable Networks
di: Juerss, Anton, et al.
Pubblicazione: (2026)
di: Juerss, Anton, et al.
Pubblicazione: (2026)
Ethereal: Divide and Conquer Network Load Balancing in Large-Scale Distributed Training
di: Addanki, Vamsi, et al.
Pubblicazione: (2024)
di: Addanki, Vamsi, et al.
Pubblicazione: (2024)
When Light Bends to the Collective Will: A Theory and Vision for Adaptive Photonic Scale-up Domains
di: Addanki, Vamsi
Pubblicazione: (2025)
di: Addanki, Vamsi
Pubblicazione: (2025)
OptiReduce: Resilient and Tail-Optimal AllReduce for Distributed Deep Learning in the Cloud
di: Warraich, Ertza, et al.
Pubblicazione: (2023)
di: Warraich, Ertza, et al.
Pubblicazione: (2023)
AllReduce Scheduling with Hierarchical Deep Reinforcement Learning
di: Wei, Yufan, et al.
Pubblicazione: (2025)
di: Wei, Yufan, et al.
Pubblicazione: (2025)
Don't Let a Few Network Failures Slow the Entire AllReduce
di: Chen, Peiqing, et al.
Pubblicazione: (2026)
di: Chen, Peiqing, et al.
Pubblicazione: (2026)
Harvest: Adaptive Photonic Switching Schedules for Collective Communication in Scale-up Domains
di: Rahman, Mahir, et al.
Pubblicazione: (2026)
di: Rahman, Mahir, et al.
Pubblicazione: (2026)
Rina: Enhancing Ring-AllReduce with In-network Aggregation in Distributed Model Training
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
Revolutionizing Datacenter Networks via Reconfigurable Topologies
di: Avin, Chen, et al.
Pubblicazione: (2025)
di: Avin, Chen, et al.
Pubblicazione: (2025)
Local Fast Rerouting with Low Congestion: A Randomized Approach
di: Bankhamer, Gregor, et al.
Pubblicazione: (2020)
di: Bankhamer, Gregor, et al.
Pubblicazione: (2020)
Risk-Aware and Stable Edge Server Selection Under Network Latency SLOs
di: Liyanage, Mohan, et al.
Pubblicazione: (2026)
di: Liyanage, Mohan, et al.
Pubblicazione: (2026)
COREC: Concurrent Non-Blocking Single-Queue Receive Driver for Low Latency Networking
di: Faltelli, Marco, et al.
Pubblicazione: (2024)
di: Faltelli, Marco, et al.
Pubblicazione: (2024)
GORGO: Maximizing KV-Cache Reuse While Minimizing Network Latency in Cross-Region LLM Load Balancing
di: Toniolo, Alessio Ricci, et al.
Pubblicazione: (2026)
di: Toniolo, Alessio Ricci, et al.
Pubblicazione: (2026)
Solving AI Foundational Model Latency with Telco Infrastructure
di: Barros, Sebastian
Pubblicazione: (2025)
di: Barros, Sebastian
Pubblicazione: (2025)
Lightweight Latency Prediction Scheme for Edge Applications: A Rational Modelling Approach
di: Liyanage, Mohan, et al.
Pubblicazione: (2025)
di: Liyanage, Mohan, et al.
Pubblicazione: (2025)
FAST: An Efficient Scheduler for All-to-All GPU Communication
di: Lei, Yiran, et al.
Pubblicazione: (2025)
di: Lei, Yiran, et al.
Pubblicazione: (2025)
RouterWise: Joint Resource Allocation and Routing for Latency-Aware Multi-Model LLM Serving
di: Kasnavieh, Hossein Hosseini, et al.
Pubblicazione: (2026)
di: Kasnavieh, Hossein Hosseini, et al.
Pubblicazione: (2026)
Efficient All-to-All Collective Communication Schedules for Direct-Connect Topologies
di: Basu, Prithwish, et al.
Pubblicazione: (2023)
di: Basu, Prithwish, et al.
Pubblicazione: (2023)
Optimal Oblivious Load-Balancing for Sparse Traffic in Large-Scale Satellite Networks
di: Ramakanth, Rudrapatna Vallabh, et al.
Pubblicazione: (2026)
di: Ramakanth, Rudrapatna Vallabh, et al.
Pubblicazione: (2026)
Dynamic Hierarchical Birkhoff-von Neumann Decomposition for All-to-All GPU Communication
di: Wu, Yen-Chieh, et al.
Pubblicazione: (2026)
di: Wu, Yen-Chieh, et al.
Pubblicazione: (2026)
RailS: Load Balancing for All-to-All Communication in Distributed Mixture-of-Experts Training
di: Xu, Heng, et al.
Pubblicazione: (2025)
di: Xu, Heng, et al.
Pubblicazione: (2025)
1.5 Million Messages Per Second on 3 Machines: Benchmarking and Latency Optimization of Apache Pulsar at Enterprise Scale
di: Mukkolakkal, Muhamed Ramees Cheriya
Pubblicazione: (2026)
di: Mukkolakkal, Muhamed Ramees Cheriya
Pubblicazione: (2026)
CRAFT: Latency and Cost-Aware Genetic-Based Framework for Node Placement in Edge-Fog Environments
di: Mahdizadeh, Soheil, et al.
Pubblicazione: (2025)
di: Mahdizadeh, Soheil, et al.
Pubblicazione: (2025)
Optimal Server Selection for Straggler Mitigation
di: Badita, Ajay, et al.
Pubblicazione: (2019)
di: Badita, Ajay, et al.
Pubblicazione: (2019)
FogROS2-PLR: Probabilistic Latency-Reliability For Cloud Robotics
di: Chen, Kaiyuan, et al.
Pubblicazione: (2024)
di: Chen, Kaiyuan, et al.
Pubblicazione: (2024)
An Auction-Based Mechanism for Optimal Task Allocation and Resource Aware Containerization
di: kumar, Ramakant
Pubblicazione: (2026)
di: kumar, Ramakant
Pubblicazione: (2026)
OptiNIC: A Resilient and Tail-Optimal RDMA NIC for Distributed ML Workloads
di: Warraich, Ertza, et al.
Pubblicazione: (2025)
di: Warraich, Ertza, et al.
Pubblicazione: (2025)
On Effectiveness of Graph Neural Network Architectures for Network Digital Twins (NDTs)
di: Zacarias, Iulisloi, et al.
Pubblicazione: (2025)
di: Zacarias, Iulisloi, et al.
Pubblicazione: (2025)
Decentralized Stratified Sampling for Low-Latency Approximate Geospatial Data Stream Processing in Edge-Cloud Architectures
di: Jawarneh, Isam Mashhour Al, et al.
Pubblicazione: (2026)
di: Jawarneh, Isam Mashhour Al, et al.
Pubblicazione: (2026)
Time Complexity of Broadcast and Consensus for Randomized Oblivious Message Adversaries
di: El-Hayek, Antoine, et al.
Pubblicazione: (2023)
di: El-Hayek, Antoine, et al.
Pubblicazione: (2023)
Extreme-Scale Interconnection Networks
di: Cano, Alejandro, et al.
Pubblicazione: (2026)
di: Cano, Alejandro, et al.
Pubblicazione: (2026)
Throughput-Optimized Networks at Scale
di: Green, Conor James, et al.
Pubblicazione: (2026)
di: Green, Conor James, et al.
Pubblicazione: (2026)
Swarm Network-as-a-Service (SNaaS)
di: Alkouz, Balsam, et al.
Pubblicazione: (2026)
di: Alkouz, Balsam, et al.
Pubblicazione: (2026)
Flooding with Absorption: An Efficient Protocol for Heterogeneous Bandits over Complex Networks
di: Lee, Junghyun, et al.
Pubblicazione: (2023)
di: Lee, Junghyun, et al.
Pubblicazione: (2023)
Runtime Verification Containers for Publish/Subscribe Networks
di: Mehran, Ali, et al.
Pubblicazione: (2024)
di: Mehran, Ali, et al.
Pubblicazione: (2024)
Placing Timely Refreshing Services at the Network Edge
di: Li, Xishuo, et al.
Pubblicazione: (2024)
di: Li, Xishuo, et al.
Pubblicazione: (2024)
SCRamble: Adaptive Decentralized Overlay Construction for Blockchain Networks
di: Kolyvas, Evangelos, et al.
Pubblicazione: (2026)
di: Kolyvas, Evangelos, et al.
Pubblicazione: (2026)
Fast Multichannel Topology Discovery in Cognitive Radio Networks
di: Wang, Yung-Li, et al.
Pubblicazione: (2025)
di: Wang, Yung-Li, et al.
Pubblicazione: (2025)
Towards Timely Video Analytics Services at the Network Edge
di: Li, Xishuo, et al.
Pubblicazione: (2024)
di: Li, Xishuo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Short-circuiting Rings for Low-Latency AllReduce
di: Hammer, Sarah-Michelle, et al.
Pubblicazione: (2025) -
Revisiting Bruck: Phase-Efficient All-to-All Communication in Reconfigurable Networks
di: Juerss, Anton, et al.
Pubblicazione: (2026) -
Ethereal: Divide and Conquer Network Load Balancing in Large-Scale Distributed Training
di: Addanki, Vamsi, et al.
Pubblicazione: (2024) -
When Light Bends to the Collective Will: A Theory and Vision for Adaptive Photonic Scale-up Domains
di: Addanki, Vamsi
Pubblicazione: (2025) -
OptiReduce: Resilient and Tail-Optimal AllReduce for Distributed Deep Learning in the Cloud
di: Warraich, Ertza, et al.
Pubblicazione: (2023)