LCI: a Lightweight Communication Interface for Efficient Asynchronous Multithreaded Communication
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yan, Jiakun, Snir, Marc |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multithreaded Fine-Grained Asynchronous BSP for Integer Sorting with LCI and OpenMP
von: Cheng, Minyu, et al.
Veröffentlicht: (2026)
von: Cheng, Minyu, et al.
Veröffentlicht: (2026)
Examining MPI and its Extensions for Asynchronous Multithreaded Communication
von: Yan, Jiakun, et al.
Veröffentlicht: (2025)
von: Yan, Jiakun, et al.
Veröffentlicht: (2025)
Contemplating a Lightweight Communication Interface for Asynchronous Many-Task Systems
von: Yan, Jiakun, et al.
Veröffentlicht: (2025)
von: Yan, Jiakun, et al.
Veröffentlicht: (2025)
Understanding the Communication Needs of Asynchronous Many-Task Systems -- A Case Study of HPX+LCI
von: Yan, Jiakun, et al.
Veröffentlicht: (2025)
von: Yan, Jiakun, et al.
Veröffentlicht: (2025)
AGILE: Lightweight and Efficient Asynchronous GPU-SSD Integration
von: Yang, Zhuoping, et al.
Veröffentlicht: (2025)
von: Yang, Zhuoping, et al.
Veröffentlicht: (2025)
Exploring Fine-grained Task Parallelism on Simultaneous Multithreading Cores
von: Los, Denis, et al.
Veröffentlicht: (2024)
von: Los, Denis, et al.
Veröffentlicht: (2024)
APEX: Asynchronous Parallel CPU-GPU Execution for Online LLM Inference on Constrained GPUs
von: Fan, Jiakun, et al.
Veröffentlicht: (2025)
von: Fan, Jiakun, et al.
Veröffentlicht: (2025)
Dynamic Simultaneous Multithreaded Architecture
von: Ortiz-Arroyo, Daniel, et al.
Veröffentlicht: (2024)
von: Ortiz-Arroyo, Daniel, et al.
Veröffentlicht: (2024)
LB4OMP: A Dynamic Load Balancing Library for Multithreaded Applications
von: Korndörfer, Jonas H. Müller, et al.
Veröffentlicht: (2021)
von: Korndörfer, Jonas H. Müller, et al.
Veröffentlicht: (2021)
Multithreaded parallelism for heterogeneous clusters of QPUs
von: Seitz, Philipp, et al.
Veröffentlicht: (2023)
von: Seitz, Philipp, et al.
Veröffentlicht: (2023)
Asynchronous Approximate Agreement with Quadratic Communication
von: Erbes, Mose Mizrahi, et al.
Veröffentlicht: (2024)
von: Erbes, Mose Mizrahi, et al.
Veröffentlicht: (2024)
Asynchronous-Many-Task Systems: Challenges and Opportunities -- Scaling an AMR Astrophysics Code on Exascale machines using Kokkos and HPX
von: Daiß, Gregor, et al.
Veröffentlicht: (2024)
von: Daiß, Gregor, et al.
Veröffentlicht: (2024)
FedADAS: Communication-Efficient Federated Distillation for On-Device Driver Yawn Recognition in Vehicular Networks
von: Mujtaba, Ahmed, et al.
Veröffentlicht: (2026)
von: Mujtaba, Ahmed, et al.
Veröffentlicht: (2026)
Formal Definitions and Performance Comparison of Consistency Models for Parallel File Systems
von: Wang, Chen, et al.
Veröffentlicht: (2024)
von: Wang, Chen, et al.
Veröffentlicht: (2024)
Communication Efficient Byzantine Agreement with Predictions
von: Dzulfikar, Muhammad Ayaz, et al.
Veröffentlicht: (2026)
von: Dzulfikar, Muhammad Ayaz, et al.
Veröffentlicht: (2026)
General Convex Agreement with Near-Optimal Communication
von: Dufay, Marc, et al.
Veröffentlicht: (2026)
von: Dufay, Marc, et al.
Veröffentlicht: (2026)
AMSP: Reducing Communication Overhead of ZeRO for Efficient LLM Training
von: Chen, Qiaoling, et al.
Veröffentlicht: (2023)
von: Chen, Qiaoling, et al.
Veröffentlicht: (2023)
Exploring the Efficiency of Renewable Energy-based Modular Data Centers at Scale
von: Sun, Jinghan, et al.
Veröffentlicht: (2024)
von: Sun, Jinghan, et al.
Veröffentlicht: (2024)
FedCod: An Efficient Communication Protocol for Cross-Silo Federated Learning with Coding
von: Yan, Peishen, et al.
Veröffentlicht: (2024)
von: Yan, Peishen, et al.
Veröffentlicht: (2024)
Recorder: Comprehensive Parallel I/O Tracing and Analysis
von: Wang, Chen, et al.
Veröffentlicht: (2025)
von: Wang, Chen, et al.
Veröffentlicht: (2025)
Computation and Communication Efficient Lightweighting Vertical Federated Learning for Smart Building IoT
von: Wang, Heqiang, et al.
Veröffentlicht: (2024)
von: Wang, Heqiang, et al.
Veröffentlicht: (2024)
Efficient Data-Parallel Continual Learning with Asynchronous Distributed Rehearsal Buffers
von: Bouvier, Thomas, et al.
Veröffentlicht: (2024)
von: Bouvier, Thomas, et al.
Veröffentlicht: (2024)
Amortized Asynchronous Byzantine Reliable Broadcast with Optimal Resilience
von: Hu, Michael Yiqing, et al.
Veröffentlicht: (2026)
von: Hu, Michael Yiqing, et al.
Veröffentlicht: (2026)
Lemonshark: Asynchronous DAG-BFT With Early Finality
von: Hu, Michael Yiqing, et al.
Veröffentlicht: (2026)
von: Hu, Michael Yiqing, et al.
Veröffentlicht: (2026)
Communication-Efficient Serving for Video Diffusion Models with Latent Parallelism
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
Enabling Scalability in Asynchronous and Bidirectional Communication in LPWAN
von: Rahman, Mahbubur
Veröffentlicht: (2025)
von: Rahman, Mahbubur
Veröffentlicht: (2025)
Taming the Memory Footprint Crisis: System Design for Production Diffusion LLM Serving
von: Fan, Jiakun, et al.
Veröffentlicht: (2025)
von: Fan, Jiakun, et al.
Veröffentlicht: (2025)
Unleashing Efficient Asynchronous RL Post-Training via Staleness-Constrained Rollout Coordination
von: Li, Haoyang, et al.
Veröffentlicht: (2026)
von: Li, Haoyang, et al.
Veröffentlicht: (2026)
An Efficient, Reliable and Observable Collective Communication Library in Large-scale GPU Training Clusters
von: Zhang, Mingjun, et al.
Veröffentlicht: (2025)
von: Zhang, Mingjun, et al.
Veröffentlicht: (2025)
Pier: Efficient Large Language Model pretraining with Relaxed Global Communication
von: Fan, Shuyuan, et al.
Veröffentlicht: (2025)
von: Fan, Shuyuan, et al.
Veröffentlicht: (2025)
Communication-Efficient Model Aggregation with Layer Divergence Feedback in Federated Learning
von: Wang, Liwei, et al.
Veröffentlicht: (2024)
von: Wang, Liwei, et al.
Veröffentlicht: (2024)
EcoFed: Efficient Communication for DNN Partitioning-based Federated Learning
von: Wu, Di, et al.
Veröffentlicht: (2023)
von: Wu, Di, et al.
Veröffentlicht: (2023)
Communication-Efficient Collaborative LLM Inference over LEO Satellite Networks
von: Zhang, Songge, et al.
Veröffentlicht: (2026)
von: Zhang, Songge, et al.
Veröffentlicht: (2026)
Communication-Efficient Distributed Learning via Sparse and Adaptive Stochastic Gradient
von: Deng, Xiaoge, et al.
Veröffentlicht: (2021)
von: Deng, Xiaoge, et al.
Veröffentlicht: (2021)
Communication-and-Computation Efficient Split Federated Learning: Gradient Aggregation and Resource Management
von: Liang, Yipeng, et al.
Veröffentlicht: (2025)
von: Liang, Yipeng, et al.
Veröffentlicht: (2025)
Communication Round and Computation Efficient Exclusive Prefix-Sums Algorithms (for MPI_Exscan)
von: Träff, Jesper Larsson
Veröffentlicht: (2025)
von: Träff, Jesper Larsson
Veröffentlicht: (2025)
Energy Efficient Federated Learning with Hyperdimensional Computing over Wireless Communication Networks
von: Ding, Yahao, et al.
Veröffentlicht: (2026)
von: Ding, Yahao, et al.
Veröffentlicht: (2026)
Efficient Local-to-Global Collaborative Perception via Joint Communication and Computation Optimization
von: Zhang, Hui, et al.
Veröffentlicht: (2026)
von: Zhang, Hui, et al.
Veröffentlicht: (2026)
Asynchronous Checkpoint for Eventually Consistent Databases
von: Ravishankar, Raaghav, et al.
Veröffentlicht: (2025)
von: Ravishankar, Raaghav, et al.
Veröffentlicht: (2025)
Byzantine Consensus in the Random Asynchronous Model
von: Danezis, George, et al.
Veröffentlicht: (2025)
von: Danezis, George, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multithreaded Fine-Grained Asynchronous BSP for Integer Sorting with LCI and OpenMP
von: Cheng, Minyu, et al.
Veröffentlicht: (2026) -
Examining MPI and its Extensions for Asynchronous Multithreaded Communication
von: Yan, Jiakun, et al.
Veröffentlicht: (2025) -
Contemplating a Lightweight Communication Interface for Asynchronous Many-Task Systems
von: Yan, Jiakun, et al.
Veröffentlicht: (2025) -
Understanding the Communication Needs of Asynchronous Many-Task Systems -- A Case Study of HPX+LCI
von: Yan, Jiakun, et al.
Veröffentlicht: (2025) -
AGILE: Lightweight and Efficient Asynchronous GPU-SSD Integration
von: Yang, Zhuoping, et al.
Veröffentlicht: (2025)