GPU-Based Parallel Computing Methods for Medical Photoacoustic Image Reconstruction
Fuente:
arXiv
Guardado en:
| Autores principales: | Yi, Xinyao, Qiao, Yuxin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Study of Performance Programming of CPU, GPU accelerated Computers and SIMD Architecture
por: Yi, Xinyao
Publicado: (2024)
por: Yi, Xinyao
Publicado: (2024)
Parallel Collaborative ADMM Privacy Computing and Adaptive GPU Acceleration for Distributed Edge Networks
por: Xia, Mengchun, et al.
Publicado: (2026)
por: Xia, Mengchun, et al.
Publicado: (2026)
Neutron particle transport 3D method of characteristic Multi GPU platform Parallel Computing
por: Zhou, Faguo, et al.
Publicado: (2025)
por: Zhou, Faguo, et al.
Publicado: (2025)
Large Scale Multi-GPU Based Parallel Traffic Simulation for Accelerated Traffic Assignment and Propagation
por: Jiang, Xuan, et al.
Publicado: (2024)
por: Jiang, Xuan, et al.
Publicado: (2024)
HARP: Orchestrating Automated Parallel Training on Heterogeneous GPU Clusters
por: Liang, Antian, et al.
Publicado: (2025)
por: Liang, Antian, et al.
Publicado: (2025)
PPipe: Efficient Video Analytics Serving on Heterogeneous GPU Clusters via Pool-Based Pipeline Parallelism
por: Kong, Z. Jonny, et al.
Publicado: (2025)
por: Kong, Z. Jonny, et al.
Publicado: (2025)
Concurrent Scheduling of High-Level Parallel Programs on Multi-GPU Systems
por: Knorr, Fabian, et al.
Publicado: (2025)
por: Knorr, Fabian, et al.
Publicado: (2025)
MERBIT: A GPU-Based SpMV Method for Iterative Workloads
por: Zhang, Qi, et al.
Publicado: (2026)
por: Zhang, Qi, et al.
Publicado: (2026)
GTaP: A GPU-Resident Fork-Join Task-Parallel Runtime with a Pragma-Based Interface
por: Maeda, Yuki, et al.
Publicado: (2026)
por: Maeda, Yuki, et al.
Publicado: (2026)
Lectures on Parallel Computing
por: Träff, Jesper Larsson
Publicado: (2024)
por: Träff, Jesper Larsson
Publicado: (2024)
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism
por: Mo, Zizhao, et al.
Publicado: (2025)
por: Mo, Zizhao, et al.
Publicado: (2025)
Multi-GPU Acceleration of PALABOS Fluid Solver using C++ Standard Parallelism
por: Latt, Jonas, et al.
Publicado: (2025)
por: Latt, Jonas, et al.
Publicado: (2025)
Heimdall++: Optimizing GPU Utilization and Pipeline Parallelism for Efficient Single-Pulse Detection
por: Xia, Bingzheng, et al.
Publicado: (2025)
por: Xia, Bingzheng, et al.
Publicado: (2025)
Parallel GPU-Enabled Algorithms for SpGEMM on Arbitrary Semirings with Hybrid Communication
por: McFarland, Thomas, et al.
Publicado: (2025)
por: McFarland, Thomas, et al.
Publicado: (2025)
EinDecomp: Decomposition of Declaratively-Specified Machine Learning and Numerical Computations for Parallel Execution
por: Bourgeois, Daniel, et al.
Publicado: (2024)
por: Bourgeois, Daniel, et al.
Publicado: (2024)
APEX: Asynchronous Parallel CPU-GPU Execution for Online LLM Inference on Constrained GPUs
por: Fan, Jiakun, et al.
Publicado: (2025)
por: Fan, Jiakun, et al.
Publicado: (2025)
Towards Compute-Aware In-Switch Computing for LLMs Tensor-Parallelism on Multi-GPU Systems
por: Zhang, Chen, et al.
Publicado: (2026)
por: Zhang, Chen, et al.
Publicado: (2026)
RAFI -- A Ray/Work Forwarding Infrastructure for Data Parallel Multi-Node/Multi-GPU Computing
por: Wald, Ingo, et al.
Publicado: (2026)
por: Wald, Ingo, et al.
Publicado: (2026)
Efficient Accelerated Graph Edit Distance Computation on GPU
por: Dabah, Adel, et al.
Publicado: (2026)
por: Dabah, Adel, et al.
Publicado: (2026)
SiPipe: Bridging the CPU-GPU Utilization Gap for Efficient Pipeline-Parallel LLM Inference
por: He, Yongchao, et al.
Publicado: (2025)
por: He, Yongchao, et al.
Publicado: (2025)
Developing an Interactive OpenMP Programming Book with Large Language Models
por: Yi, Xinyao, et al.
Publicado: (2024)
por: Yi, Xinyao, et al.
Publicado: (2024)
Parallel Watershed Partitioning: GPU-Based Hierarchical Image Segmentation
por: Yeghiazaryan, Varduhi, et al.
Publicado: (2024)
por: Yeghiazaryan, Varduhi, et al.
Publicado: (2024)
EXaCTz: Guaranteed Extremum Graph and Contour Tree Preservation for Distributed- and GPU-Parallel Lossy Compression
por: Li, Yuxiao, et al.
Publicado: (2026)
por: Li, Yuxiao, et al.
Publicado: (2026)
Towards Fast Setup and High Throughput of GPU Serverless Computing
por: Zhao, Han, et al.
Publicado: (2024)
por: Zhao, Han, et al.
Publicado: (2024)
GRNND: A GPU-Parallel Relative NN-Descent Algorithm for Efficient Approximate Nearest Neighbor Graph Construction
por: Li, Xiang, et al.
Publicado: (2025)
por: Li, Xiang, et al.
Publicado: (2025)
Evaluation of Programming Models and Performance for Stencil Computation on Current GPU Architectures
por: Shan, Baodi, et al.
Publicado: (2024)
por: Shan, Baodi, et al.
Publicado: (2024)
FaaSTube: Optimizing GPU-oriented Data Transfer for Serverless Computing
por: Wu, Hao, et al.
Publicado: (2024)
por: Wu, Hao, et al.
Publicado: (2024)
cuFastTuckerPlus: A Stochastic Parallel Sparse FastTucker Decomposition Using GPU Tensor Cores
por: Li, Zixuan, et al.
Publicado: (2024)
por: Li, Zixuan, et al.
Publicado: (2024)
Warp-STAR: High-performance, Differentiable GPU-Accelerated Static Timing Analysis through Warp-oriented Parallel Orchestration
por: Huang, En-Ming, et al.
Publicado: (2026)
por: Huang, En-Ming, et al.
Publicado: (2026)
A Unified CPU-GPU Protocol for GNN Training
por: Lin, Yi-Chien, et al.
Publicado: (2024)
por: Lin, Yi-Chien, et al.
Publicado: (2024)
TURNIP: A "Nondeterministic" GPU Runtime with CPU RAM Offload
por: Ding, Zhimin, et al.
Publicado: (2024)
por: Ding, Zhimin, et al.
Publicado: (2024)
An AD based library for Efficient Hessian and Hessian-Vector Product Computation on GPU
por: Ranjan, Desh, et al.
Publicado: (2024)
por: Ranjan, Desh, et al.
Publicado: (2024)
Torpor: GPU-Enabled Serverless Computing for Low-Latency, Resource-Efficient Inference
por: Yu, Minchen, et al.
Publicado: (2023)
por: Yu, Minchen, et al.
Publicado: (2023)
ZEUS: An Efficient GPU Optimization Method Integrating PSO, BFGS, and Automatic Differentiation
por: Soos, Dominik, et al.
Publicado: (2026)
por: Soos, Dominik, et al.
Publicado: (2026)
Fantasy: Efficient Large-scale Vector Search on GPU Clusters with GPUDirect Async
por: Liu, Yi, et al.
Publicado: (2025)
por: Liu, Yi, et al.
Publicado: (2025)
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
por: Qiao, Tong, et al.
Publicado: (2025)
por: Qiao, Tong, et al.
Publicado: (2025)
CusADi: A GPU Parallelization Framework for Symbolic Expressions and Optimal Control
por: Jeon, Se Hwan, et al.
Publicado: (2024)
por: Jeon, Se Hwan, et al.
Publicado: (2024)
DuaLip-GPU Technical Report
por: Dexter, Gregory, et al.
Publicado: (2026)
por: Dexter, Gregory, et al.
Publicado: (2026)
Enhancing ASIC Technology Mapping via Parallel Supergate Computing
por: Cai, Ye, et al.
Publicado: (2024)
por: Cai, Ye, et al.
Publicado: (2024)
Resource-efficient Parallel Split Learning in Heterogeneous Edge Computing
por: Zhang, Mingjin, et al.
Publicado: (2024)
por: Zhang, Mingjin, et al.
Publicado: (2024)
Ejemplares similares
-
A Study of Performance Programming of CPU, GPU accelerated Computers and SIMD Architecture
por: Yi, Xinyao
Publicado: (2024) -
Parallel Collaborative ADMM Privacy Computing and Adaptive GPU Acceleration for Distributed Edge Networks
por: Xia, Mengchun, et al.
Publicado: (2026) -
Neutron particle transport 3D method of characteristic Multi GPU platform Parallel Computing
por: Zhou, Faguo, et al.
Publicado: (2025) -
Large Scale Multi-GPU Based Parallel Traffic Simulation for Accelerated Traffic Assignment and Propagation
por: Jiang, Xuan, et al.
Publicado: (2024) -
HARP: Orchestrating Automated Parallel Training on Heterogeneous GPU Clusters
por: Liang, Antian, et al.
Publicado: (2025)