MPI Errors Detection using GNN Embedding and Vector Embedding over LLVM IR
Fuente:
arXiv
Salvato in:
| Autori principali: | Karchi, Jad El, Chen, Hanze, TehraniJamsaz, Ali, Jannesari, Ali, Popov, Mihail, Saillard, Emmanuelle |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CodeRosetta: Pushing the Boundaries of Unsupervised Code Translation for Parallel Programming
di: TehraniJamsaz, Ali, et al.
Pubblicazione: (2024)
di: TehraniJamsaz, Ali, et al.
Pubblicazione: (2024)
OMPILOT: Harnessing Transformer Models for Auto Parallelization to Shared Memory Computing Paradigms
di: Bhattacharjee, Arijit, et al.
Pubblicazione: (2025)
di: Bhattacharjee, Arijit, et al.
Pubblicazione: (2025)
Enabling MPI communication within Numba/LLVM JIT-compiled Python code using numba-mpi v1.0
di: Derlatka, Kacper, et al.
Pubblicazione: (2024)
di: Derlatka, Kacper, et al.
Pubblicazione: (2024)
MIREncoder: Multi-modal IR-based Pretrained Embeddings for Performance Optimizations
di: Dutta, Akash, et al.
Pubblicazione: (2024)
di: Dutta, Akash, et al.
Pubblicazione: (2024)
An LLVM-Based Optimization Pipeline for SPDZ
di: Dai, Tianye, et al.
Pubblicazione: (2025)
di: Dai, Tianye, et al.
Pubblicazione: (2025)
NOMAD: Generating Embeddings for Massive Distributed Graphs
di: Sarkar, Aishwarya, et al.
Pubblicazione: (2026)
di: Sarkar, Aishwarya, et al.
Pubblicazione: (2026)
MPI Implementation Profiling for Better Application Performance
di: Shipley, Riley, et al.
Pubblicazione: (2024)
di: Shipley, Riley, et al.
Pubblicazione: (2024)
Performance measurements of modern Fortran MPI applications with Score-P
di: Corbin, Gregor
Pubblicazione: (2025)
di: Corbin, Gregor
Pubblicazione: (2025)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
di: Diehl, Patrick, et al.
Pubblicazione: (2025)
di: Diehl, Patrick, et al.
Pubblicazione: (2025)
Automated MPI-X code generation for scalable finite-difference solvers
di: Bisbas, George, et al.
Pubblicazione: (2023)
di: Bisbas, George, et al.
Pubblicazione: (2023)
Incremental GNN Embedding Computation on Streaming Graphs
di: Wang, Qiange, et al.
Pubblicazione: (2026)
di: Wang, Qiange, et al.
Pubblicazione: (2026)
OMPGPT: A Generative Pre-trained Transformer Model for OpenMP
di: Chen, Le, et al.
Pubblicazione: (2024)
di: Chen, Le, et al.
Pubblicazione: (2024)
Dynamic Detection of Inefficient Data Mapping Patterns in Heterogeneous OpenMP Applications
di: Marzen, Luke, et al.
Pubblicazione: (2026)
di: Marzen, Luke, et al.
Pubblicazione: (2026)
Static Generation of Efficient OpenMP Offload Data Mappings
di: Marzen, Luke, et al.
Pubblicazione: (2024)
di: Marzen, Luke, et al.
Pubblicazione: (2024)
MassiveGNN: Efficient Training via Prefetching for Massively Connected Distributed Graphs
di: Sarkar, Aishwarya, et al.
Pubblicazione: (2024)
di: Sarkar, Aishwarya, et al.
Pubblicazione: (2024)
WindVE: Collaborative CPU-NPU Vector Embedding
di: Huang, Jinqi, et al.
Pubblicazione: (2025)
di: Huang, Jinqi, et al.
Pubblicazione: (2025)
PennyLane-Lightning MPI: A massively scalable quantum circuit simulator based on distributed computing in CPU clusters
di: Kang, Ji-Hoon, et al.
Pubblicazione: (2025)
di: Kang, Ji-Hoon, et al.
Pubblicazione: (2025)
An Optimized Error-controlled MPI Collective Framework Integrated with Lossy Compression
di: Huang, Jiajun, et al.
Pubblicazione: (2023)
di: Huang, Jiajun, et al.
Pubblicazione: (2023)
Comprehensive Review of Performance Optimization Strategies for Serverless Applications on AWS Lambda
di: Bechir, Mohamed Lemine El, et al.
Pubblicazione: (2024)
di: Bechir, Mohamed Lemine El, et al.
Pubblicazione: (2024)
Implementing True MPI Sessions and Evaluating MPI Initialization Scalability
di: Zhou, Hui, et al.
Pubblicazione: (2026)
di: Zhou, Hui, et al.
Pubblicazione: (2026)
MPI-over-CXL: Enhancing Communication Efficiency in Distributed HPC Systems
di: Kwon, Miryeong, et al.
Pubblicazione: (2025)
di: Kwon, Miryeong, et al.
Pubblicazione: (2025)
MPI Progress For All
di: Zhou, Hui, et al.
Pubblicazione: (2024)
di: Zhou, Hui, et al.
Pubblicazione: (2024)
Towards Secure Management of Edge-Cloud IoT Microservices using Policy as Code
di: Pallewatta, Samodha, et al.
Pubblicazione: (2024)
di: Pallewatta, Samodha, et al.
Pubblicazione: (2024)
SuperSFL: Resource-Heterogeneous Federated Split Learning with Weight-Sharing Super-Networks
di: Asif, Abdullah Al, et al.
Pubblicazione: (2026)
di: Asif, Abdullah Al, et al.
Pubblicazione: (2026)
LibProf: A Python Profiler for Improving Cold Start Performance in Serverless Applications
di: Tariq, Syed Salauddin Mohammad, et al.
Pubblicazione: (2024)
di: Tariq, Syed Salauddin Mohammad, et al.
Pubblicazione: (2024)
Scaling MPI Applications on Aurora
di: Ibeid, Huda, et al.
Pubblicazione: (2025)
di: Ibeid, Huda, et al.
Pubblicazione: (2025)
On the performance of two-sided MPI, MPI-3 RMA and SHMEM in a Lagrangian particle cluster algorithm
di: Frey, Matthias, et al.
Pubblicazione: (2024)
di: Frey, Matthias, et al.
Pubblicazione: (2024)
Persistent and Partitioned MPI for Stencil Communication
di: Collom, Gerald, et al.
Pubblicazione: (2025)
di: Collom, Gerald, et al.
Pubblicazione: (2025)
Some New Approaches to MPI Implementations
di: Xiong, Yuqing
Pubblicazione: (2024)
di: Xiong, Yuqing
Pubblicazione: (2024)
Designing and Prototyping Extensions to MPI in MPICH
di: Zhou, Hui, et al.
Pubblicazione: (2024)
di: Zhou, Hui, et al.
Pubblicazione: (2024)
HeLoCo: Efficient asynchronous low-communication training under data and device heterogeneity
di: Asif, Abdullah Al, et al.
Pubblicazione: (2026)
di: Asif, Abdullah Al, et al.
Pubblicazione: (2026)
Embedded Made Easy -- Rethinking Embedded + Cloud Software Development (WIP)
di: Arnold, Anthony, et al.
Pubblicazione: (2026)
di: Arnold, Anthony, et al.
Pubblicazione: (2026)
OptiML: An End-to-End Framework for Program Synthesis and CUDA Kernel Optimization
di: Bhattacharjee, Arijit, et al.
Pubblicazione: (2026)
di: Bhattacharjee, Arijit, et al.
Pubblicazione: (2026)
Concepts for designing modern C++ interfaces for MPI
di: Avans, C. Nicole, et al.
Pubblicazione: (2025)
di: Avans, C. Nicole, et al.
Pubblicazione: (2025)
Frustrated with MPI+Threads? Try MPIxThreads!
di: Zhou, Hui, et al.
Pubblicazione: (2024)
di: Zhou, Hui, et al.
Pubblicazione: (2024)
Interferences within a certifiable design methodology for high-performance multi-core platforms
di: Khelassi, Mohamed Amine, et al.
Pubblicazione: (2026)
di: Khelassi, Mohamed Amine, et al.
Pubblicazione: (2026)
Towards Lock Modularization for Heterogeneous Environments
di: Zhang, Hanze, et al.
Pubblicazione: (2025)
di: Zhang, Hanze, et al.
Pubblicazione: (2025)
Data Race Satisfiability on Array Elements
di: Shim, Junhyung, et al.
Pubblicazione: (2025)
di: Shim, Junhyung, et al.
Pubblicazione: (2025)
SeBS-Flow: Benchmarking Serverless Cloud Function Workflows
di: Schmid, Larissa, et al.
Pubblicazione: (2024)
di: Schmid, Larissa, et al.
Pubblicazione: (2024)
A Unifying Framework to Enable Artificial Intelligence in High Performance Computing Workflows
di: Domke, Jens, et al.
Pubblicazione: (2025)
di: Domke, Jens, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CodeRosetta: Pushing the Boundaries of Unsupervised Code Translation for Parallel Programming
di: TehraniJamsaz, Ali, et al.
Pubblicazione: (2024) -
OMPILOT: Harnessing Transformer Models for Auto Parallelization to Shared Memory Computing Paradigms
di: Bhattacharjee, Arijit, et al.
Pubblicazione: (2025) -
Enabling MPI communication within Numba/LLVM JIT-compiled Python code using numba-mpi v1.0
di: Derlatka, Kacper, et al.
Pubblicazione: (2024) -
MIREncoder: Multi-modal IR-based Pretrained Embeddings for Performance Optimizations
di: Dutta, Akash, et al.
Pubblicazione: (2024) -
An LLVM-Based Optimization Pipeline for SPDZ
di: Dai, Tianye, et al.
Pubblicazione: (2025)