Parallel CPU- and GPU-based connected component algorithms for event building for hybrid pixel detectors
Fuente:
arXiv
Guardado en:
| Autores principales: | Čelko, Tomáš, Mráz, František, Bergmann, Benedikt, Mánek, Petr |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Track Lab: extensible data acquisition software for fast pixel detectors, online analysis and automation
por: Mánek, Petr, et al.
Publicado: (2023)
por: Mánek, Petr, et al.
Publicado: (2023)
SuperSONIC: Cloud-Native Infrastructure for ML Inferencing
por: Kondratyev, Dmitry, et al.
Publicado: (2025)
por: Kondratyev, Dmitry, et al.
Publicado: (2025)
Comparative Analysis of FPGA and GPU Performance for Machine Learning-Based Track Reconstruction at LHCb
por: Giasemis, Fotis I., et al.
Publicado: (2025)
por: Giasemis, Fotis I., et al.
Publicado: (2025)
Portable acceleration of CMS computing workflows with coprocessors as a service
por: CMS Collaboration
Publicado: (2024)
por: CMS Collaboration
Publicado: (2024)
Track reconstruction as a service for collider physics
por: Zhao, Haoran, et al.
Publicado: (2025)
por: Zhao, Haoran, et al.
Publicado: (2025)
Advancing ATLAS DCS Data Analysis with a Modern Data Platform
por: Canali, Luca, et al.
Publicado: (2025)
por: Canali, Luca, et al.
Publicado: (2025)
APEX: Asynchronous Parallel CPU-GPU Execution for Online LLM Inference on Constrained GPUs
por: Fan, Jiakun, et al.
Publicado: (2025)
por: Fan, Jiakun, et al.
Publicado: (2025)
Streaming Large-Scale Electron Microscopy Data to a Supercomputing Facility
por: Welborn, Samuel S., et al.
Publicado: (2024)
por: Welborn, Samuel S., et al.
Publicado: (2024)
SiPipe: Bridging the CPU-GPU Utilization Gap for Efficient Pipeline-Parallel LLM Inference
por: He, Yongchao, et al.
Publicado: (2025)
por: He, Yongchao, et al.
Publicado: (2025)
A Unified CPU-GPU Protocol for GNN Training
por: Lin, Yi-Chien, et al.
Publicado: (2024)
por: Lin, Yi-Chien, et al.
Publicado: (2024)
Combining GPU and CPU for accelerating evolutionary computing workloads
por: Eynaliyev, Rustam, et al.
Publicado: (2025)
por: Eynaliyev, Rustam, et al.
Publicado: (2025)
TURNIP: A "Nondeterministic" GPU Runtime with CPU RAM Offload
por: Ding, Zhimin, et al.
Publicado: (2024)
por: Ding, Zhimin, et al.
Publicado: (2024)
Efficient CPU-GPU Collaborative Inference for MoE-based LLMs on Memory-Limited Systems
por: Huang, En-Ming, et al.
Publicado: (2025)
por: Huang, En-Ming, et al.
Publicado: (2025)
Efficient Graph Embedding at Scale: Optimizing CPU-GPU-SSD Integration
por: Li, Zhonggen, et al.
Publicado: (2025)
por: Li, Zhonggen, et al.
Publicado: (2025)
Towards Affordable, Adaptive and Automatic GNN Training on CPU-GPU Heterogeneous Platforms
por: Qiao, Tong, et al.
Publicado: (2025)
por: Qiao, Tong, et al.
Publicado: (2025)
Taming GPU Underutilization via Static Partitioning and Fine-grained CPU Offloading
por: Schieffer, Gabin, et al.
Publicado: (2026)
por: Schieffer, Gabin, et al.
Publicado: (2026)
A Study of Performance Programming of CPU, GPU accelerated Computers and SIMD Architecture
por: Yi, Xinyao
Publicado: (2024)
por: Yi, Xinyao
Publicado: (2024)
A Parallel CPU-GPU Framework for Batching Heuristic Operations in Depth-First Heuristic Search
por: Futuhi, Ehsan, et al.
Publicado: (2025)
por: Futuhi, Ehsan, et al.
Publicado: (2025)
votess: A multi-target, GPU-capable, parallel Voronoi tessellator
por: Singh, Samridh Dev, et al.
Publicado: (2024)
por: Singh, Samridh Dev, et al.
Publicado: (2024)
Co-Design and Evaluation of a CPU-Free MPI GPU Communication Abstraction and Implementation
por: Bridges, Patrick G., et al.
Publicado: (2026)
por: Bridges, Patrick G., et al.
Publicado: (2026)
Serving Hybrid LLM Loads with SLO Guarantees Using CPU-GPU Attention Piggybacking
por: Mo, Zizhao, et al.
Publicado: (2026)
por: Mo, Zizhao, et al.
Publicado: (2026)
EuroHPC SPACE CoE: Redesigning Scalable Parallel Astrophysical Codes for Exascale
por: Shukla, Nitin, et al.
Publicado: (2025)
por: Shukla, Nitin, et al.
Publicado: (2025)
Breaking the Memory Wall: A Study of I/O Patterns and GPU Memory Utilization for Hybrid CPU-GPU Offloaded Optimizers
por: Maurya, Avinash, et al.
Publicado: (2024)
por: Maurya, Avinash, et al.
Publicado: (2024)
Harnessing Integrated CPU-GPU System Memory for HPC: a first look into Grace Hopper
por: Schieffer, Gabin, et al.
Publicado: (2024)
por: Schieffer, Gabin, et al.
Publicado: (2024)
Cost-Performance Analysis: A Comparative Study of CPU-Based Serverless and GPU-Based Training Architectures
por: Barrak, Amine, et al.
Publicado: (2025)
por: Barrak, Amine, et al.
Publicado: (2025)
Evaluation of computational and energy performance in matrix multiplication algorithms on CPU and GPU using MKL, cuBLAS and SYCL
por: Torres, L. A., et al.
Publicado: (2024)
por: Torres, L. A., et al.
Publicado: (2024)
HeteroSTA: A CPU-GPU Heterogeneous Static Timing Analysis Engine with Holistic Industrial Design Support
por: Guo, Zizheng, et al.
Publicado: (2025)
por: Guo, Zizheng, et al.
Publicado: (2025)
Orchestrated Co-scheduling, Resource Partitioning, and Power Capping on CPU-GPU Heterogeneous Systems via Machine Learning
por: Saba, Issa, et al.
Publicado: (2024)
por: Saba, Issa, et al.
Publicado: (2024)
Dissecting CPU-GPU Unified Physical Memory on AMD MI300A APUs
por: Wahlgren, Jacob, et al.
Publicado: (2025)
por: Wahlgren, Jacob, et al.
Publicado: (2025)
HybridGen: Efficient LLM Generative Inference via CPU-GPU Hybrid Computing
por: Lin, Mao, et al.
Publicado: (2026)
por: Lin, Mao, et al.
Publicado: (2026)
Characterizing CPU-Induced Slowdowns in Multi-GPU LLM Inference
por: Chung, Euijun, et al.
Publicado: (2026)
por: Chung, Euijun, et al.
Publicado: (2026)
Unified schemes for directive-based GPU offloading
por: Miki, Yohei, et al.
Publicado: (2024)
por: Miki, Yohei, et al.
Publicado: (2024)
HARP: Orchestrating Automated Parallel Training on Heterogeneous GPU Clusters
por: Liang, Antian, et al.
Publicado: (2025)
por: Liang, Antian, et al.
Publicado: (2025)
Workflow decomposition algorithm for scheduling with quantum annealer-based hybrid solver
por: Kroczek, Marcin, et al.
Publicado: (2025)
por: Kroczek, Marcin, et al.
Publicado: (2025)
GPU-Based Parallel Computing Methods for Medical Photoacoustic Image Reconstruction
por: Yi, Xinyao, et al.
Publicado: (2024)
por: Yi, Xinyao, et al.
Publicado: (2024)
Concurrent Scheduling of High-Level Parallel Programs on Multi-GPU Systems
por: Knorr, Fabian, et al.
Publicado: (2025)
por: Knorr, Fabian, et al.
Publicado: (2025)
Towards CXL Resilience to CPU Failures
por: Psistakis, Antonis, et al.
Publicado: (2026)
por: Psistakis, Antonis, et al.
Publicado: (2026)
Hetis: Serving LLMs in Heterogeneous GPU Clusters with Fine-grained and Dynamic Parallelism
por: Mo, Zizhao, et al.
Publicado: (2025)
por: Mo, Zizhao, et al.
Publicado: (2025)
Multi-GPU Acceleration of PALABOS Fluid Solver using C++ Standard Parallelism
por: Latt, Jonas, et al.
Publicado: (2025)
por: Latt, Jonas, et al.
Publicado: (2025)
Heimdall++: Optimizing GPU Utilization and Pipeline Parallelism for Efficient Single-Pulse Detection
por: Xia, Bingzheng, et al.
Publicado: (2025)
por: Xia, Bingzheng, et al.
Publicado: (2025)
Ejemplares similares
-
Track Lab: extensible data acquisition software for fast pixel detectors, online analysis and automation
por: Mánek, Petr, et al.
Publicado: (2023) -
SuperSONIC: Cloud-Native Infrastructure for ML Inferencing
por: Kondratyev, Dmitry, et al.
Publicado: (2025) -
Comparative Analysis of FPGA and GPU Performance for Machine Learning-Based Track Reconstruction at LHCb
por: Giasemis, Fotis I., et al.
Publicado: (2025) -
Portable acceleration of CMS computing workflows with coprocessors as a service
por: CMS Collaboration
Publicado: (2024) -
Track reconstruction as a service for collider physics
por: Zhao, Haoran, et al.
Publicado: (2025)