Short-circuiting Rings for Low-Latency AllReduce
Fuente:
arXiv
Salvato in:
| Autori principali: | Hammer, Sarah-Michelle, Schmid, Stefan, Singh, Rachee, Addanki, Vamsi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Trivance: Latency-Optimal AllReduce by Shortcutting Multiport Networks
di: Juerss, Anton, et al.
Pubblicazione: (2026)
di: Juerss, Anton, et al.
Pubblicazione: (2026)
When Light Bends to the Collective Will: A Theory and Vision for Adaptive Photonic Scale-up Domains
di: Addanki, Vamsi
Pubblicazione: (2025)
di: Addanki, Vamsi
Pubblicazione: (2025)
Ethereal: Divide and Conquer Network Load Balancing in Large-Scale Distributed Training
di: Addanki, Vamsi, et al.
Pubblicazione: (2024)
di: Addanki, Vamsi, et al.
Pubblicazione: (2024)
AllReduce Scheduling with Hierarchical Deep Reinforcement Learning
di: Wei, Yufan, et al.
Pubblicazione: (2025)
di: Wei, Yufan, et al.
Pubblicazione: (2025)
OptiReduce: Resilient and Tail-Optimal AllReduce for Distributed Deep Learning in the Cloud
di: Warraich, Ertza, et al.
Pubblicazione: (2023)
di: Warraich, Ertza, et al.
Pubblicazione: (2023)
Rina: Enhancing Ring-AllReduce with In-network Aggregation in Distributed Model Training
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
Harvest: Adaptive Photonic Switching Schedules for Collective Communication in Scale-up Domains
di: Rahman, Mahir, et al.
Pubblicazione: (2026)
di: Rahman, Mahir, et al.
Pubblicazione: (2026)
Don't Let a Few Network Failures Slow the Entire AllReduce
di: Chen, Peiqing, et al.
Pubblicazione: (2026)
di: Chen, Peiqing, et al.
Pubblicazione: (2026)
Revisiting Bruck: Phase-Efficient All-to-All Communication in Reconfigurable Networks
di: Juerss, Anton, et al.
Pubblicazione: (2026)
di: Juerss, Anton, et al.
Pubblicazione: (2026)
Local Fast Rerouting with Low Congestion: A Randomized Approach
di: Bankhamer, Gregor, et al.
Pubblicazione: (2020)
di: Bankhamer, Gregor, et al.
Pubblicazione: (2020)
Revolutionizing Datacenter Networks via Reconfigurable Topologies
di: Avin, Chen, et al.
Pubblicazione: (2025)
di: Avin, Chen, et al.
Pubblicazione: (2025)
COREC: Concurrent Non-Blocking Single-Queue Receive Driver for Low Latency Networking
di: Faltelli, Marco, et al.
Pubblicazione: (2024)
di: Faltelli, Marco, et al.
Pubblicazione: (2024)
CCL-Bench 1.0: A Trace-Based Benchmark for LLM Infrastructure
di: Ding, Eric, et al.
Pubblicazione: (2026)
di: Ding, Eric, et al.
Pubblicazione: (2026)
Solving AI Foundational Model Latency with Telco Infrastructure
di: Barros, Sebastian
Pubblicazione: (2025)
di: Barros, Sebastian
Pubblicazione: (2025)
Efficient AllReduce with Stragglers
di: Devraj, Arjun, et al.
Pubblicazione: (2025)
di: Devraj, Arjun, et al.
Pubblicazione: (2025)
Lightweight Latency Prediction Scheme for Edge Applications: A Rational Modelling Approach
di: Liyanage, Mohan, et al.
Pubblicazione: (2025)
di: Liyanage, Mohan, et al.
Pubblicazione: (2025)
Risk-Aware and Stable Edge Server Selection Under Network Latency SLOs
di: Liyanage, Mohan, et al.
Pubblicazione: (2026)
di: Liyanage, Mohan, et al.
Pubblicazione: (2026)
FAST: An Efficient Scheduler for All-to-All GPU Communication
di: Lei, Yiran, et al.
Pubblicazione: (2025)
di: Lei, Yiran, et al.
Pubblicazione: (2025)
RouterWise: Joint Resource Allocation and Routing for Latency-Aware Multi-Model LLM Serving
di: Kasnavieh, Hossein Hosseini, et al.
Pubblicazione: (2026)
di: Kasnavieh, Hossein Hosseini, et al.
Pubblicazione: (2026)
Efficient All-to-All Collective Communication Schedules for Direct-Connect Topologies
di: Basu, Prithwish, et al.
Pubblicazione: (2023)
di: Basu, Prithwish, et al.
Pubblicazione: (2023)
Decentralized Stratified Sampling for Low-Latency Approximate Geospatial Data Stream Processing in Edge-Cloud Architectures
di: Jawarneh, Isam Mashhour Al, et al.
Pubblicazione: (2026)
di: Jawarneh, Isam Mashhour Al, et al.
Pubblicazione: (2026)
GORGO: Maximizing KV-Cache Reuse While Minimizing Network Latency in Cross-Region LLM Load Balancing
di: Toniolo, Alessio Ricci, et al.
Pubblicazione: (2026)
di: Toniolo, Alessio Ricci, et al.
Pubblicazione: (2026)
RailS: Load Balancing for All-to-All Communication in Distributed Mixture-of-Experts Training
di: Xu, Heng, et al.
Pubblicazione: (2025)
di: Xu, Heng, et al.
Pubblicazione: (2025)
Dynamic Hierarchical Birkhoff-von Neumann Decomposition for All-to-All GPU Communication
di: Wu, Yen-Chieh, et al.
Pubblicazione: (2026)
di: Wu, Yen-Chieh, et al.
Pubblicazione: (2026)
1.5 Million Messages Per Second on 3 Machines: Benchmarking and Latency Optimization of Apache Pulsar at Enterprise Scale
di: Mukkolakkal, Muhamed Ramees Cheriya
Pubblicazione: (2026)
di: Mukkolakkal, Muhamed Ramees Cheriya
Pubblicazione: (2026)
CRAFT: Latency and Cost-Aware Genetic-Based Framework for Node Placement in Edge-Fog Environments
di: Mahdizadeh, Soheil, et al.
Pubblicazione: (2025)
di: Mahdizadeh, Soheil, et al.
Pubblicazione: (2025)
FogROS2-PLR: Probabilistic Latency-Reliability For Cloud Robotics
di: Chen, Kaiyuan, et al.
Pubblicazione: (2024)
di: Chen, Kaiyuan, et al.
Pubblicazione: (2024)
Hurry: Dynamic Collaborative Framework For Low-orbit Mega-Constellation Data Downloading
di: Luo, Handong, et al.
Pubblicazione: (2024)
di: Luo, Handong, et al.
Pubblicazione: (2024)
Edge-assisted Parallel Uncertain Skyline Processing for Low-latency IoE Analysis
di: Lai, Chuan-Chi, et al.
Pubblicazione: (2025)
di: Lai, Chuan-Chi, et al.
Pubblicazione: (2025)
LOAM: Low-latency Communication, Caching, and Computation Placement in Data-Intensive Computing Networks
di: Zhang, Jinkun, et al.
Pubblicazione: (2024)
di: Zhang, Jinkun, et al.
Pubblicazione: (2024)
Time Complexity of Broadcast and Consensus for Randomized Oblivious Message Adversaries
di: El-Hayek, Antoine, et al.
Pubblicazione: (2023)
di: El-Hayek, Antoine, et al.
Pubblicazione: (2023)
DRDST: Low-latency DAG Consensus through Robust Dynamic Sharding and Tree-broadcasting for IoV
di: Chen, Runhua, et al.
Pubblicazione: (2024)
di: Chen, Runhua, et al.
Pubblicazione: (2024)
AirChain: A Novel Blockchain Framework and Low-Cost Device for Democratized Air Quality Data Aggregation
di: Stankiewicz, Samuel
Pubblicazione: (2024)
di: Stankiewicz, Samuel
Pubblicazione: (2024)
POSMAC: Powering Up In-Network AR/CG Traffic Classification with Online Learning
di: Shirmarz, Alireza, et al.
Pubblicazione: (2025)
di: Shirmarz, Alireza, et al.
Pubblicazione: (2025)
Low-Latency Video Conferencing via Optimized Packet Routing and Reordering
di: Xiao, Yao, et al.
Pubblicazione: (2023)
di: Xiao, Yao, et al.
Pubblicazione: (2023)
Analysing Mechanisms for Virtual Channel Management in Low-Diameter networks
di: Cano, Alejandro, et al.
Pubblicazione: (2023)
di: Cano, Alejandro, et al.
Pubblicazione: (2023)
A New Broadcast Primitive for BFT Protocols
di: Drijvers, Manu, et al.
Pubblicazione: (2024)
di: Drijvers, Manu, et al.
Pubblicazione: (2024)
Optimizing Split Learning Latency in TinyML-Based IoT Systems
di: Jenhani, Zied, et al.
Pubblicazione: (2025)
di: Jenhani, Zied, et al.
Pubblicazione: (2025)
RailX: A Flexible, Scalable, and Low-Cost Network Architecture for Hyper-Scale LLM Training Systems
di: Feng, Yinxiao, et al.
Pubblicazione: (2025)
di: Feng, Yinxiao, et al.
Pubblicazione: (2025)
On the Resilience of Fast Failover Routing Against Dynamic Link Failures
di: Dai, Wenkai, et al.
Pubblicazione: (2024)
di: Dai, Wenkai, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Trivance: Latency-Optimal AllReduce by Shortcutting Multiport Networks
di: Juerss, Anton, et al.
Pubblicazione: (2026) -
When Light Bends to the Collective Will: A Theory and Vision for Adaptive Photonic Scale-up Domains
di: Addanki, Vamsi
Pubblicazione: (2025) -
Ethereal: Divide and Conquer Network Load Balancing in Large-Scale Distributed Training
di: Addanki, Vamsi, et al.
Pubblicazione: (2024) -
AllReduce Scheduling with Hierarchical Deep Reinforcement Learning
di: Wei, Yufan, et al.
Pubblicazione: (2025) -
OptiReduce: Resilient and Tail-Optimal AllReduce for Distributed Deep Learning in the Cloud
di: Warraich, Ertza, et al.
Pubblicazione: (2023)