Sharded Elimination and Combining for Highly-Efficient Concurrent Stacks
Fuente:
arXiv
Salvato in:
| Autori principali: | Singh, Ajay, Metaxakis, Nikos, Fatourou, Panagiota |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
NVLang: Unified Static Typing for Actor-Based Concurrency on the BEAM
di: Guerreiro, Miguel de Oliveira
Pubblicazione: (2025)
di: Guerreiro, Miguel de Oliveira
Pubblicazione: (2025)
Highly-Efficient Persistent FIFO Queues
di: Fatourou, Panagiota, et al.
Pubblicazione: (2024)
di: Fatourou, Panagiota, et al.
Pubblicazione: (2024)
VeriFx: Correct Replicated Data Types for the Masses
di: De Porre, Kevin, et al.
Pubblicazione: (2022)
di: De Porre, Kevin, et al.
Pubblicazione: (2022)
Fancy Some Chips for Your TeaStore? Modeling the Control of an Adaptable Discrete System
di: Gallone, Anna, et al.
Pubblicazione: (2025)
di: Gallone, Anna, et al.
Pubblicazione: (2025)
AutoTSMM: An Auto-tuning Framework for Building High-Performance Tall-and-Skinny Matrix-Matrix Multiplication on CPUs
di: Li, Chendi, et al.
Pubblicazione: (2022)
di: Li, Chendi, et al.
Pubblicazione: (2022)
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
di: Kanakagiri, Raghavendra, et al.
Pubblicazione: (2023)
di: Kanakagiri, Raghavendra, et al.
Pubblicazione: (2023)
Categorical Message Passing Language (CaMPL) for programmers
di: Hashimoto, Daniel Kiyoshi, et al.
Pubblicazione: (2026)
di: Hashimoto, Daniel Kiyoshi, et al.
Pubblicazione: (2026)
SpaDA: A Spatial Dataflow Architecture Programming Language
di: Gianinazzi, Lukas, et al.
Pubblicazione: (2025)
di: Gianinazzi, Lukas, et al.
Pubblicazione: (2025)
Optimizing Fine-Grained Parallelism Through Dynamic Load Balancing on Multi-Socket Many-Core Systems
di: Wang, Wenyi, et al.
Pubblicazione: (2025)
di: Wang, Wenyi, et al.
Pubblicazione: (2025)
Enabling Practical Transparent Checkpointing for MPI: A Topological Sort Approach
di: Xu, Yao, et al.
Pubblicazione: (2024)
di: Xu, Yao, et al.
Pubblicazione: (2024)
Publish on Ping: A Better Way to Publish Reservations in Memory Reclamation for Concurrent Data Structures
di: Singh, Ajay, et al.
Pubblicazione: (2025)
di: Singh, Ajay, et al.
Pubblicazione: (2025)
A C++17 Thread Pool for High-Performance Scientific Computing
di: Shoshany, Barak
Pubblicazione: (2021)
di: Shoshany, Barak
Pubblicazione: (2021)
A Formal Semantics of C with OpenMP Parallelism (Extended Version)
di: Du, Ke, et al.
Pubblicazione: (2026)
di: Du, Ke, et al.
Pubblicazione: (2026)
Challenging Portability Paradigms: FPGA Acceleration Using SYCL and OpenCL
di: de Castro, Manuel, et al.
Pubblicazione: (2024)
di: de Castro, Manuel, et al.
Pubblicazione: (2024)
Stream parallel skeleton optimization
di: Aldinucci, Marco, et al.
Pubblicazione: (2024)
di: Aldinucci, Marco, et al.
Pubblicazione: (2024)
StreamFlow: cross-breeding cloud with HPC
di: Colonnelli, Iacopo, et al.
Pubblicazione: (2020)
di: Colonnelli, Iacopo, et al.
Pubblicazione: (2020)
How to Relax Instantly: Elastic Relaxation of Concurrent Data Structures
di: von Geijer, Kåre, et al.
Pubblicazione: (2024)
di: von Geijer, Kåre, et al.
Pubblicazione: (2024)
HTVM: Efficient Neural Network Deployment On Heterogeneous TinyML Platforms
di: Van Delm, Josse, et al.
Pubblicazione: (2024)
di: Van Delm, Josse, et al.
Pubblicazione: (2024)
DNA sequence alignment: An assignment for OpenMP, MPI, and CUDA/OpenCL
di: Gonzalez-Escribano, Arturo, et al.
Pubblicazione: (2024)
di: Gonzalez-Escribano, Arturo, et al.
Pubblicazione: (2024)
Scalable Concurrent Queues for GPU
di: Shetty, Pratheek Prakash, et al.
Pubblicazione: (2026)
di: Shetty, Pratheek Prakash, et al.
Pubblicazione: (2026)
Dynamic Memory Management on GPUs with SYCL
di: Standish, Russell K.
Pubblicazione: (2025)
di: Standish, Russell K.
Pubblicazione: (2025)
StableShard: Stable and Scalable Blockchain Sharding with High Concurrency via Collaborative Committees
di: Li, Mingzhe, et al.
Pubblicazione: (2024)
di: Li, Mingzhe, et al.
Pubblicazione: (2024)
Hybrid Quantum-HPC Middleware Systems for Adaptive Resource, Workload and Task Management
di: Mantha, Pradeep, et al.
Pubblicazione: (2026)
di: Mantha, Pradeep, et al.
Pubblicazione: (2026)
Lock-Free Augmented Trees
di: Fatourou, Panagiota, et al.
Pubblicazione: (2024)
di: Fatourou, Panagiota, et al.
Pubblicazione: (2024)
Recoverable Lock-Free Locks
di: Attiya, Hagit, et al.
Pubblicazione: (2025)
di: Attiya, Hagit, et al.
Pubblicazione: (2025)
Utilizing Sparsity in the GPU-accelerated Assembly of Schur Complement Matrices in Domain Decomposition Methods
di: Homola, Jakub, et al.
Pubblicazione: (2025)
di: Homola, Jakub, et al.
Pubblicazione: (2025)
Modular GPU Programming with Typed Perspectives
di: Bansal, Manya, et al.
Pubblicazione: (2025)
di: Bansal, Manya, et al.
Pubblicazione: (2025)
Construction of a Byzantine Linearizable SWMR Atomic Register from SWSR Atomic Registers
di: Kshemkalyani, Ajay D., et al.
Pubblicazione: (2024)
di: Kshemkalyani, Ajay D., et al.
Pubblicazione: (2024)
Concurrent Data Structures Made Easy (Extended Version)
di: Le, Callista, et al.
Pubblicazione: (2024)
di: Le, Callista, et al.
Pubblicazione: (2024)
Joint Training on AMD and NVIDIA GPUs
di: Hu, Jon, et al.
Pubblicazione: (2026)
di: Hu, Jon, et al.
Pubblicazione: (2026)
Static Batching of Irregular Workloads on GPUs: Framework and Application to Efficient MoE Model Inference
di: Li, Yinghan, et al.
Pubblicazione: (2025)
di: Li, Yinghan, et al.
Pubblicazione: (2025)
HiCoCS: High Concurrency Cross-Sharding on Permissioned Blockchains
di: Yang, Lingxiao, et al.
Pubblicazione: (2025)
di: Yang, Lingxiao, et al.
Pubblicazione: (2025)
The Autonomous Data Language -- Concepts, Design and Formal Verification
di: Franken, Tom T. P., et al.
Pubblicazione: (2025)
di: Franken, Tom T. P., et al.
Pubblicazione: (2025)
VSS Challenge Problem: Verifying the Correctness of AllReduce Algorithms in the MPICH Implementation of MPI
di: Hovland, Paul D.
Pubblicazione: (2025)
di: Hovland, Paul D.
Pubblicazione: (2025)
An Evaluation of Massively Parallel Algorithms for DFA Minimization
di: Martens, Jan, et al.
Pubblicazione: (2024)
di: Martens, Jan, et al.
Pubblicazione: (2024)
Abstract Continuation Semantics for Multiparty Interactions in Process Calculi based on CCS
di: Todoran, Eneia Nicolae, et al.
Pubblicazione: (2024)
di: Todoran, Eneia Nicolae, et al.
Pubblicazione: (2024)
NM-SpMM: Accelerating Matrix Multiplication Using N:M Sparsity with GPGPU
di: Ma, Cong, et al.
Pubblicazione: (2025)
di: Ma, Cong, et al.
Pubblicazione: (2025)
Hector: An Efficient Programming and Compilation Framework for Implementing Relational Graph Neural Networks in GPU Architectures
di: Wu, Kun, et al.
Pubblicazione: (2023)
di: Wu, Kun, et al.
Pubblicazione: (2023)
Two-sorted algebraic decompositions of Brookes's shared-state denotational semantics
di: Dvir, Yotam, et al.
Pubblicazione: (2025)
di: Dvir, Yotam, et al.
Pubblicazione: (2025)
Introducing SWIRL: An Intermediate Representation Language for Scientific Workflows
di: Colonnelli, Iacopo, et al.
Pubblicazione: (2024)
di: Colonnelli, Iacopo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
NVLang: Unified Static Typing for Actor-Based Concurrency on the BEAM
di: Guerreiro, Miguel de Oliveira
Pubblicazione: (2025) -
Highly-Efficient Persistent FIFO Queues
di: Fatourou, Panagiota, et al.
Pubblicazione: (2024) -
VeriFx: Correct Replicated Data Types for the Masses
di: De Porre, Kevin, et al.
Pubblicazione: (2022) -
Fancy Some Chips for Your TeaStore? Modeling the Control of an Adaptable Discrete System
di: Gallone, Anna, et al.
Pubblicazione: (2025) -
AutoTSMM: An Auto-tuning Framework for Building High-Performance Tall-and-Skinny Matrix-Matrix Multiplication on CPUs
di: Li, Chendi, et al.
Pubblicazione: (2022)