HiRace: Accurate and Fast Source-Level Race Checking of GPU Programs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jacobson, John, Burtscher, Martin, Gopalakrishnan, Ganesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SIMT/GPU Data Race Verification using ISCC and Intermediary Code Representations: A Case Study
von: Osterhout, Andrew, et al.
Veröffentlicht: (2025)
von: Osterhout, Andrew, et al.
Veröffentlicht: (2025)
Fast Topology-Aware Lossy Data Compression with Full Preservation of Critical Points and Local Order
von: Fallin, Alex, et al.
Veröffentlicht: (2026)
von: Fallin, Alex, et al.
Veröffentlicht: (2026)
Lessons Learned on the Path to Guaranteeing the Error Bound in Lossy Quantizers
von: Fallin, Alex, et al.
Veröffentlicht: (2024)
von: Fallin, Alex, et al.
Veröffentlicht: (2024)
A GPU accelerated mixed-precision Smoothed Particle Hydrodynamics framework with cell-based relative coordinates
von: Mao, Zirui, et al.
Veröffentlicht: (2023)
von: Mao, Zirui, et al.
Veröffentlicht: (2023)
HiCR, an Abstract Model for Distributed Heterogeneous Programming
von: Martin, Sergio Miguel, et al.
Veröffentlicht: (2025)
von: Martin, Sergio Miguel, et al.
Veröffentlicht: (2025)
Concurrent Scheduling of High-Level Parallel Programs on Multi-GPU Systems
von: Knorr, Fabian, et al.
Veröffentlicht: (2025)
von: Knorr, Fabian, et al.
Veröffentlicht: (2025)
HMTRace: Hardware-Assisted Memory-Tagging based Dynamic Data Race Detection
von: Shastri, Jaidev, et al.
Veröffentlicht: (2024)
von: Shastri, Jaidev, et al.
Veröffentlicht: (2024)
FastGraph: Optimized GPU-Enabled Algorithms for Fast Graph Building and Message Passing
von: Agarwal, Aarush, et al.
Veröffentlicht: (2025)
von: Agarwal, Aarush, et al.
Veröffentlicht: (2025)
Understanding GPU Resource Interference One Level Deeper
von: Elvinger, Paul, et al.
Veröffentlicht: (2025)
von: Elvinger, Paul, et al.
Veröffentlicht: (2025)
Towards Fast Setup and High Throughput of GPU Serverless Computing
von: Zhao, Han, et al.
Veröffentlicht: (2024)
von: Zhao, Han, et al.
Veröffentlicht: (2024)
VDCores: Resource Decoupled Programming and Execution for Asynchronous GPU
von: He, Zijian, et al.
Veröffentlicht: (2026)
von: He, Zijian, et al.
Veröffentlicht: (2026)
Data Race Satisfiability on Array Elements
von: Shim, Junhyung, et al.
Veröffentlicht: (2025)
von: Shim, Junhyung, et al.
Veröffentlicht: (2025)
Evaluation of Programming Models and Performance for Stencil Computation on Current GPU Architectures
von: Shan, Baodi, et al.
Veröffentlicht: (2024)
von: Shan, Baodi, et al.
Veröffentlicht: (2024)
cuFastTuckerPlus: A Stochastic Parallel Sparse FastTucker Decomposition Using GPU Tensor Cores
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
von: Li, Zixuan, et al.
Veröffentlicht: (2024)
TurboFFT: A High-Performance Fast Fourier Transform with Fault Tolerance on GPU
von: Wu, Shixun, et al.
Veröffentlicht: (2024)
von: Wu, Shixun, et al.
Veröffentlicht: (2024)
HAP: SPMD DNN Training on Heterogeneous GPU Clusters with Automated Program Synthesis
von: Zhang, Shiwei, et al.
Veröffentlicht: (2024)
von: Zhang, Shiwei, et al.
Veröffentlicht: (2024)
A Study of Performance Programming of CPU, GPU accelerated Computers and SIMD Architecture
von: Yi, Xinyao
Veröffentlicht: (2024)
von: Yi, Xinyao
Veröffentlicht: (2024)
GPU Programming for AI Workflow Development on AWS SageMaker: An Instructional Approach
von: Srinivasan, Sriram, et al.
Veröffentlicht: (2025)
von: Srinivasan, Sriram, et al.
Veröffentlicht: (2025)
AGAThA: Fast and Efficient GPU Acceleration of Guided Sequence Alignment for Long Read Mapping
von: Park, Seongyeon, et al.
Veröffentlicht: (2024)
von: Park, Seongyeon, et al.
Veröffentlicht: (2024)
FastTrack: GPU-Accelerated Tracking for Visual SLAM
von: Khabiri, Kimia, et al.
Veröffentlicht: (2025)
von: Khabiri, Kimia, et al.
Veröffentlicht: (2025)
TopoSZp: Lightweight Topology-Aware Error-controlled Compression for Scientific Data
von: Agarwal, Tripti, et al.
Veröffentlicht: (2026)
von: Agarwal, Tripti, et al.
Veröffentlicht: (2026)
Taking GPU Programming Models to Task for Performance Portability
von: Davis, Joshua H., et al.
Veröffentlicht: (2024)
von: Davis, Joshua H., et al.
Veröffentlicht: (2024)
HiCCL: A Hierarchical Collective Communication Library
von: Hidayetoglu, Mert, et al.
Veröffentlicht: (2024)
von: Hidayetoglu, Mert, et al.
Veröffentlicht: (2024)
Cppless: Single-Source and High-Performance Serverless Programming in C++
von: Copik, Marcin, et al.
Veröffentlicht: (2024)
von: Copik, Marcin, et al.
Veröffentlicht: (2024)
Improving GPU Multi-Tenancy Through Dynamic Multi-Instance GPU Reconfiguration
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
HoSZp: An Efficient Homomorphic Error-bounded Lossy Compressor for Scientific Data
von: Agarwal, Tripti, et al.
Veröffentlicht: (2024)
von: Agarwal, Tripti, et al.
Veröffentlicht: (2024)
HiCoCS: High Concurrency Cross-Sharding on Permissioned Blockchains
von: Yang, Lingxiao, et al.
Veröffentlicht: (2025)
von: Yang, Lingxiao, et al.
Veröffentlicht: (2025)
Accelerating Biclique Counting on GPU
von: Qiu, Linshan, et al.
Veröffentlicht: (2024)
von: Qiu, Linshan, et al.
Veröffentlicht: (2024)
GPU Sharing with Triples Mode
von: Byun, Chansup, et al.
Veröffentlicht: (2024)
von: Byun, Chansup, et al.
Veröffentlicht: (2024)
ParvaGPU: Efficient Spatial GPU Sharing for Large-Scale DNN Inference in Cloud Environments
von: Lee, Munkyu, et al.
Veröffentlicht: (2024)
von: Lee, Munkyu, et al.
Veröffentlicht: (2024)
Accelerating Intra-Node GPU-to-GPU Communication Through Multi-Path Transfers with CUDA Graphs
von: Sojoodi, Amirhossein, et al.
Veröffentlicht: (2026)
von: Sojoodi, Amirhossein, et al.
Veröffentlicht: (2026)
Fast and Scalable Mixed Precision Euclidean Distance Calculations Using GPU Tensor Cores
von: Curless, Brian, et al.
Veröffentlicht: (2025)
von: Curless, Brian, et al.
Veröffentlicht: (2025)
GPU Accelerated Sparse Cholesky Factorization
von: Karsavuran, M. Ozan, et al.
Veröffentlicht: (2024)
von: Karsavuran, M. Ozan, et al.
Veröffentlicht: (2024)
Heat: Satellite's meat is GPU's poison
von: Yuan, Zhehu, et al.
Veröffentlicht: (2024)
von: Yuan, Zhehu, et al.
Veröffentlicht: (2024)
DuaLip-GPU Technical Report
von: Dexter, Gregory, et al.
Veröffentlicht: (2026)
von: Dexter, Gregory, et al.
Veröffentlicht: (2026)
Incidence Constraints in Hypergraph Partitioning on GPU
von: Ronzani, Marco, et al.
Veröffentlicht: (2026)
von: Ronzani, Marco, et al.
Veröffentlicht: (2026)
Predictable LLM Serving on GPU Clusters
von: Darzi, Erfan, et al.
Veröffentlicht: (2025)
von: Darzi, Erfan, et al.
Veröffentlicht: (2025)
CheckMate: LLM-Powered Approximate Intermittent Computing
von: Sayyid-Ali, Abdur-Rahman Ibrahim, et al.
Veröffentlicht: (2024)
von: Sayyid-Ali, Abdur-Rahman Ibrahim, et al.
Veröffentlicht: (2024)
HAS-GPU: Efficient Hybrid Auto-scaling with Fine-grained GPU Allocation for SLO-aware Serverless Inferences
von: Gu, Jianfeng, et al.
Veröffentlicht: (2025)
von: Gu, Jianfeng, et al.
Veröffentlicht: (2025)
HiRL: Hierarchical Reinforcement Learning for Coordinated Resource Management in Heterogeneous Edge Computing
von: Zhu, Jianyong, et al.
Veröffentlicht: (2026)
von: Zhu, Jianyong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SIMT/GPU Data Race Verification using ISCC and Intermediary Code Representations: A Case Study
von: Osterhout, Andrew, et al.
Veröffentlicht: (2025) -
Fast Topology-Aware Lossy Data Compression with Full Preservation of Critical Points and Local Order
von: Fallin, Alex, et al.
Veröffentlicht: (2026) -
Lessons Learned on the Path to Guaranteeing the Error Bound in Lossy Quantizers
von: Fallin, Alex, et al.
Veröffentlicht: (2024) -
A GPU accelerated mixed-precision Smoothed Particle Hydrodynamics framework with cell-based relative coordinates
von: Mao, Zirui, et al.
Veröffentlicht: (2023) -
HiCR, an Abstract Model for Distributed Heterogeneous Programming
von: Martin, Sergio Miguel, et al.
Veröffentlicht: (2025)