Combining Performance and Productivity: Accelerating the Network Sensing Graph Challenge with GPUs and Commodity Data Science Software
Fuente:
arXiv
Salvato in:
| Autori principali: | Samsi, Siddharth, Campbell, Dan, Scoullos, Emanuel, Green, Oded |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Accelerating Maximal Biclique Enumeration on GPUs
di: Hsieh, Chou-Ying, et al.
Pubblicazione: (2024)
di: Hsieh, Chou-Ying, et al.
Pubblicazione: (2024)
Anonymized Network Sensing using C++26 std::execution on GPUs
di: Mandulak, Michael, et al.
Pubblicazione: (2025)
di: Mandulak, Michael, et al.
Pubblicazione: (2025)
Boosting Performance of Iterative Applications on GPUs: Kernel Batching with CUDA Graphs
di: Ekelund, Jonah, et al.
Pubblicazione: (2025)
di: Ekelund, Jonah, et al.
Pubblicazione: (2025)
RTop-K: Ultra-Fast Row-Wise Top-K Selection for Neural Network Acceleration on GPUs
di: Xie, Xi, et al.
Pubblicazione: (2024)
di: Xie, Xi, et al.
Pubblicazione: (2024)
Accelerating Sparse Matrix-Matrix Multiplication on GPUs with Processing Near HBMs
di: Li, Shiju, et al.
Pubblicazione: (2025)
di: Li, Shiju, et al.
Pubblicazione: (2025)
AMPED: Accelerating MTTKRP for Billion-Scale Sparse Tensor Decomposition on Multiple GPUs
di: Wijeratne, Sasindu, et al.
Pubblicazione: (2025)
di: Wijeratne, Sasindu, et al.
Pubblicazione: (2025)
Popcorn: Accelerating Kernel K-means on GPUs through Sparse Linear Algebra
di: Bellavita, Julian, et al.
Pubblicazione: (2025)
di: Bellavita, Julian, et al.
Pubblicazione: (2025)
Accelerating high-order continuum kinetic plasma simulations using multiple GPUs
di: Ho, Andrew, et al.
Pubblicazione: (2024)
di: Ho, Andrew, et al.
Pubblicazione: (2024)
TrioSeq: A Novel Approach to Accelerate Triplet Sequence Alignment on GPUs
di: Graça, Miguel, et al.
Pubblicazione: (2026)
di: Graça, Miguel, et al.
Pubblicazione: (2026)
Analytical Performance Estimation during Code Generation on Modern GPUs
di: Ernst, Dominik, et al.
Pubblicazione: (2022)
di: Ernst, Dominik, et al.
Pubblicazione: (2022)
XaaS: Acceleration as a Service to Enable Productive High-Performance Cloud Computing
di: Hoefler, Torsten, et al.
Pubblicazione: (2024)
di: Hoefler, Torsten, et al.
Pubblicazione: (2024)
Cuckoo-GPU: Accelerating Cuckoo Filters on Modern GPUs
di: Dortmann, Tim, et al.
Pubblicazione: (2026)
di: Dortmann, Tim, et al.
Pubblicazione: (2026)
RL over Commodity Networks: Overcoming the Bandwidth Barrier with Lossless Sparse Deltas
di: Ruan, Chaoyi, et al.
Pubblicazione: (2026)
di: Ruan, Chaoyi, et al.
Pubblicazione: (2026)
Performance Portable Monte Carlo Particle Transport on Intel, NVIDIA, and AMD GPUs
di: Tramm, John, et al.
Pubblicazione: (2024)
di: Tramm, John, et al.
Pubblicazione: (2024)
LuWu: An End-to-End In-Network Out-of-Core Optimizer for 100B-Scale Model-in-Network Data-Parallel Training on Distributed GPUs
di: Sun, Mo, et al.
Pubblicazione: (2024)
di: Sun, Mo, et al.
Pubblicazione: (2024)
BOA Constrictor: Squeezing Performance out of GPUs in the Cloud via Budget-Optimal Allocation
di: Li, Zhouzi, et al.
Pubblicazione: (2026)
di: Li, Zhouzi, et al.
Pubblicazione: (2026)
Enabling Seamless Data Security, Consensus, and Trading in Vehicular Networks
di: Vieira, Emanuel, et al.
Pubblicazione: (2024)
di: Vieira, Emanuel, et al.
Pubblicazione: (2024)
Analyzing the Performance Portability of SYCL across CPUs, GPUs, and Hybrid Systems with SW Sequence Alignment
di: Costanzo, Manuel, et al.
Pubblicazione: (2024)
di: Costanzo, Manuel, et al.
Pubblicazione: (2024)
TurboFFT: Co-Designed High-Performance and Fault-Tolerant Fast Fourier Transform on GPUs
di: Wu, Shixun, et al.
Pubblicazione: (2024)
di: Wu, Shixun, et al.
Pubblicazione: (2024)
HP-MDR: High-performance and Portable Data Refactoring and Progressive Retrieval with Advanced GPUs
di: Li, Yanliang, et al.
Pubblicazione: (2025)
di: Li, Yanliang, et al.
Pubblicazione: (2025)
An Adaptive Distributed Stencil Abstraction for GPUs
di: Bhosale, Aditya, et al.
Pubblicazione: (2025)
di: Bhosale, Aditya, et al.
Pubblicazione: (2025)
Parallelizing Maximal Clique Enumeration on GPUs
di: Almasri, Mohammad, et al.
Pubblicazione: (2022)
di: Almasri, Mohammad, et al.
Pubblicazione: (2022)
Optimizing sDTW for AMD GPUs
di: Latta-Lin, Daniel, et al.
Pubblicazione: (2024)
di: Latta-Lin, Daniel, et al.
Pubblicazione: (2024)
Performance Evaluation of Hashing Algorithms on Commodity Hardware
di: Pandya, Marut
Pubblicazione: (2024)
di: Pandya, Marut
Pubblicazione: (2024)
HetCCL: Accelerating LLM Training with Heterogeneous GPUs
di: Kim, Heehoon, et al.
Pubblicazione: (2026)
di: Kim, Heehoon, et al.
Pubblicazione: (2026)
Serving Compound Inference Systems on Datacenter GPUs
di: Devata, Sriram, et al.
Pubblicazione: (2026)
di: Devata, Sriram, et al.
Pubblicazione: (2026)
Fast Kronecker Matrix-Matrix Multiplication on GPUs
di: Jangda, Abhinav, et al.
Pubblicazione: (2024)
di: Jangda, Abhinav, et al.
Pubblicazione: (2024)
Optimal Workload Placement on Multi-Instance GPUs
di: Turkkan, Bekir, et al.
Pubblicazione: (2024)
di: Turkkan, Bekir, et al.
Pubblicazione: (2024)
Scalable Graph Indexing using GPUs for Approximate Nearest Neighbor Search
di: Li, Zhonggen, et al.
Pubblicazione: (2025)
di: Li, Zhonggen, et al.
Pubblicazione: (2025)
LLMQ: Efficient Lower-Precision Pretraining for Consumer GPUs
di: Schultheis, Erik, et al.
Pubblicazione: (2025)
di: Schultheis, Erik, et al.
Pubblicazione: (2025)
PipeMax: Enhancing Offline LLM Inference on Commodity GPU Servers
di: Zhang, Hongbin, et al.
Pubblicazione: (2026)
di: Zhang, Hongbin, et al.
Pubblicazione: (2026)
GROMACS Unplugged: How Power Capping and Frequency Shapes Performance on GPUs
di: Afzal, Ayesha, et al.
Pubblicazione: (2025)
di: Afzal, Ayesha, et al.
Pubblicazione: (2025)
Straggler Tolerant and Resilient DL Training on Homogeneous GPUs
di: Zhang, Zeyu, et al.
Pubblicazione: (2025)
di: Zhang, Zeyu, et al.
Pubblicazione: (2025)
RDMA-Based Algorithms for Sparse Matrix Multiplication on GPUs
di: Brock, Benjamin, et al.
Pubblicazione: (2023)
di: Brock, Benjamin, et al.
Pubblicazione: (2023)
Accurate Computation of the Logarithm of Modified Bessel Functions on GPUs
di: Plesner, Andreas, et al.
Pubblicazione: (2024)
di: Plesner, Andreas, et al.
Pubblicazione: (2024)
Accelerating MoE with Dynamic In-Switch Computing on Multi-GPUs
di: Zhang, Qijun, et al.
Pubblicazione: (2026)
di: Zhang, Qijun, et al.
Pubblicazione: (2026)
AcceleratedLiNGAM: Learning Causal DAGs at the speed of GPUs
di: Akinwande, Victor, et al.
Pubblicazione: (2024)
di: Akinwande, Victor, et al.
Pubblicazione: (2024)
Accelerating the Delivery of Data Services over Uncertain Mobile Crowdsensing Networks
di: Liwang, Minghui, et al.
Pubblicazione: (2022)
di: Liwang, Minghui, et al.
Pubblicazione: (2022)
Managing Multi Instance GPUs for High Throughput and Energy Savings
di: Saraha, Abhijeet, et al.
Pubblicazione: (2025)
di: Saraha, Abhijeet, et al.
Pubblicazione: (2025)
Demystifying Cost-Efficiency in LLM Serving over Heterogeneous GPUs
di: Jiang, Youhe, et al.
Pubblicazione: (2025)
di: Jiang, Youhe, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Accelerating Maximal Biclique Enumeration on GPUs
di: Hsieh, Chou-Ying, et al.
Pubblicazione: (2024) -
Anonymized Network Sensing using C++26 std::execution on GPUs
di: Mandulak, Michael, et al.
Pubblicazione: (2025) -
Boosting Performance of Iterative Applications on GPUs: Kernel Batching with CUDA Graphs
di: Ekelund, Jonah, et al.
Pubblicazione: (2025) -
RTop-K: Ultra-Fast Row-Wise Top-K Selection for Neural Network Acceleration on GPUs
di: Xie, Xi, et al.
Pubblicazione: (2024) -
Accelerating Sparse Matrix-Matrix Multiplication on GPUs with Processing Near HBMs
di: Li, Shiju, et al.
Pubblicazione: (2025)