Pragma driven shared memory parallelism in Zig by supporting OpenMP loop directives
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kacs, David, Lee, Joseph, Zarins, Justs, Brown, Nick |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Implementing OpenMP for Zig to enable its use in HPC context
von: Kacs, David, et al.
Veröffentlicht: (2024)
von: Kacs, David, et al.
Veröffentlicht: (2024)
An MLIR pipeline for offloading Fortran to FPGAs via OpenMP
von: Rodriguez-Canal, Gabriel, et al.
Veröffentlicht: (2025)
von: Rodriguez-Canal, Gabriel, et al.
Veröffentlicht: (2025)
An MLIR Lowering Pipeline for Stencils at Wafer-Scale
von: Stawinoga, Nicolai, et al.
Veröffentlicht: (2026)
von: Stawinoga, Nicolai, et al.
Veröffentlicht: (2026)
Detrimental task execution patterns in mainstream OpenMP runtimes
von: Tuft, Adam S., et al.
Veröffentlicht: (2024)
von: Tuft, Adam S., et al.
Veröffentlicht: (2024)
OMP4Py: a pure Python implementation of OpenMP
von: Piñeiro, César, et al.
Veröffentlicht: (2024)
von: Piñeiro, César, et al.
Veröffentlicht: (2024)
A Formal Semantics of C with OpenMP Parallelism (Extended Version)
von: Du, Ke, et al.
Veröffentlicht: (2026)
von: Du, Ke, et al.
Veröffentlicht: (2026)
Fully integrating the Flang Fortran compiler with standard MLIR
von: Brown, Nick
Veröffentlicht: (2024)
von: Brown, Nick
Veröffentlicht: (2024)
Static Generation of Efficient OpenMP Offload Data Mappings
von: Marzen, Luke, et al.
Veröffentlicht: (2024)
von: Marzen, Luke, et al.
Veröffentlicht: (2024)
Enabling performance portability of data-parallel OpenMP applications on asymmetric multicore processors
von: Saez, Juan Carlos, et al.
Veröffentlicht: (2024)
von: Saez, Juan Carlos, et al.
Veröffentlicht: (2024)
Distributed OpenMP Offloading of OpenMC on Intel GPU MAX Accelerators
von: Fridman, Yehonatan, et al.
Veröffentlicht: (2024)
von: Fridman, Yehonatan, et al.
Veröffentlicht: (2024)
DiOMP-Offloading: Toward Portable Distributed Heterogeneous OpenMP
von: Shan, Baodi, et al.
Veröffentlicht: (2025)
von: Shan, Baodi, et al.
Veröffentlicht: (2025)
Developing an Interactive OpenMP Programming Book with Large Language Models
von: Yi, Xinyao, et al.
Veröffentlicht: (2024)
von: Yi, Xinyao, et al.
Veröffentlicht: (2024)
LLOR: Automated Repair of OpenMP Programs
von: Bora, Utpal, et al.
Veröffentlicht: (2024)
von: Bora, Utpal, et al.
Veröffentlicht: (2024)
Dynamic Detection of Inefficient Data Mapping Patterns in Heterogeneous OpenMP Applications
von: Marzen, Luke, et al.
Veröffentlicht: (2026)
von: Marzen, Luke, et al.
Veröffentlicht: (2026)
Auto-Tuning for OpenMP Dynamic Scheduling applied to Full Waveform Inversion
von: da Silva, Felipe H. S., et al.
Veröffentlicht: (2024)
von: da Silva, Felipe H. S., et al.
Veröffentlicht: (2024)
Multithreaded Fine-Grained Asynchronous BSP for Integer Sorting with LCI and OpenMP
von: Cheng, Minyu, et al.
Veröffentlicht: (2026)
von: Cheng, Minyu, et al.
Veröffentlicht: (2026)
Accelerating Fortran Codes: A Method for Integrating Coarray Fortran with CUDA Fortran and OpenMP
von: McKevitt, James, et al.
Veröffentlicht: (2024)
von: McKevitt, James, et al.
Veröffentlicht: (2024)
Parallel Paradigms in Modern HPC: A Comparative Analysis of MPI, OpenMP, and CUDA
von: ALHafez, Nizar, et al.
Veröffentlicht: (2025)
von: ALHafez, Nizar, et al.
Veröffentlicht: (2025)
Towards a Scalable and Efficient PGAS-based Distributed OpenMP
von: Shan, Baodi, et al.
Veröffentlicht: (2024)
von: Shan, Baodi, et al.
Veröffentlicht: (2024)
OMP-Engineer: Bridging Syntax Analysis and In-Context Learning for Efficient Automated OpenMP Parallelization
von: Wang, Weidong, et al.
Veröffentlicht: (2024)
von: Wang, Weidong, et al.
Veröffentlicht: (2024)
Enhanced OpenMP Algorithm to Compute All-Pairs Shortest Path on x86 Architectures
von: Calderón, Sergio, et al.
Veröffentlicht: (2024)
von: Calderón, Sergio, et al.
Veröffentlicht: (2024)
A shared compilation stack for distributed-memory parallelism in stencil DSLs
von: Bisbas, George, et al.
Veröffentlicht: (2024)
von: Bisbas, George, et al.
Veröffentlicht: (2024)
Porting HPC Applications to AMD Instinct$^\text{TM}$ MI300A Using Unified Memory and OpenMP
von: Tandon, Suyash, et al.
Veröffentlicht: (2024)
von: Tandon, Suyash, et al.
Veröffentlicht: (2024)
Publish on Ping: A Better Way to Publish Reservations in Memory Reclamation for Concurrent Data Structures
von: Singh, Ajay, et al.
Veröffentlicht: (2025)
von: Singh, Ajay, et al.
Veröffentlicht: (2025)
High-Performance Parallel Optimization of the Fish School Behaviour on the Setonix Platform Using OpenMP
von: Wang, Haitian, et al.
Veröffentlicht: (2025)
von: Wang, Haitian, et al.
Veröffentlicht: (2025)
Parallel FFTW on RISC-V: A Comparative Study including OpenMP, MPI, and HPX
von: Strack, Alexander, et al.
Veröffentlicht: (2025)
von: Strack, Alexander, et al.
Veröffentlicht: (2025)
Scaling Sample-Based Quantum Diagonalization on GPU-Accelerated Systems using OpenMP Offload
von: Walkup, Robert, et al.
Veröffentlicht: (2026)
von: Walkup, Robert, et al.
Veröffentlicht: (2026)
OMPGPT: A Generative Pre-trained Transformer Model for OpenMP
von: Chen, Le, et al.
Veröffentlicht: (2024)
von: Chen, Le, et al.
Veröffentlicht: (2024)
Optimizing the Weather Research and Forecasting Model with OpenMP Offload and Codee
von: Chayanon, et al.
Veröffentlicht: (2024)
von: Chayanon, et al.
Veröffentlicht: (2024)
We Know I Know You Know; Choreographic Programming With Multicast and Multiply Located Values
von: Bates, Mako, et al.
Veröffentlicht: (2024)
von: Bates, Mako, et al.
Veröffentlicht: (2024)
Suki: Choreographed Distributed Dataflow in Rust
von: Laddad, Shadaj, et al.
Veröffentlicht: (2024)
von: Laddad, Shadaj, et al.
Veröffentlicht: (2024)
MultiChor: Census Polymorphic Choreographic Programming with Multiply Located Values
von: Bates, Mako, et al.
Veröffentlicht: (2024)
von: Bates, Mako, et al.
Veröffentlicht: (2024)
KPerfIR: Towards an Open and Compiler-centric Ecosystem for GPU Kernel Performance Tooling on Modern AI Workloads
von: Guan, Yue, et al.
Veröffentlicht: (2025)
von: Guan, Yue, et al.
Veröffentlicht: (2025)
Flo: a Semantic Foundation for Progressive Stream Processing
von: Laddad, Shadaj, et al.
Veröffentlicht: (2024)
von: Laddad, Shadaj, et al.
Veröffentlicht: (2024)
Accelerating Particle-in-Cell Monte Carlo Simulations with MPI, OpenMP/OpenACC and Asynchronous Multi-GPU Programming
von: Williams, Jeremy J., et al.
Veröffentlicht: (2024)
von: Williams, Jeremy J., et al.
Veröffentlicht: (2024)
CPU-less parallel execution of lambda calculus in digital logic
von: Fitchett, Harry, et al.
Veröffentlicht: (2026)
von: Fitchett, Harry, et al.
Veröffentlicht: (2026)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
von: Diehl, Patrick, et al.
Veröffentlicht: (2025)
von: Diehl, Patrick, et al.
Veröffentlicht: (2025)
A Comparative Study of OpenMP Scheduling Algorithm Selection Strategies
von: Korndörfer, Jonas H. Müller, et al.
Veröffentlicht: (2025)
von: Korndörfer, Jonas H. Müller, et al.
Veröffentlicht: (2025)
On the Duality of Task and Actor Programming Models
von: Yadav, Rohan, et al.
Veröffentlicht: (2025)
von: Yadav, Rohan, et al.
Veröffentlicht: (2025)
Concurrent Data Structures Made Easy (Extended Version)
von: Le, Callista, et al.
Veröffentlicht: (2024)
von: Le, Callista, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Implementing OpenMP for Zig to enable its use in HPC context
von: Kacs, David, et al.
Veröffentlicht: (2024) -
An MLIR pipeline for offloading Fortran to FPGAs via OpenMP
von: Rodriguez-Canal, Gabriel, et al.
Veröffentlicht: (2025) -
An MLIR Lowering Pipeline for Stencils at Wafer-Scale
von: Stawinoga, Nicolai, et al.
Veröffentlicht: (2026) -
Detrimental task execution patterns in mainstream OpenMP runtimes
von: Tuft, Adam S., et al.
Veröffentlicht: (2024) -
OMP4Py: a pure Python implementation of OpenMP
von: Piñeiro, César, et al.
Veröffentlicht: (2024)