FastGraph: Optimized GPU-Enabled Algorithms for Fast Graph Building and Message Passing
Fuente:
arXiv
Saved in:
| Main Authors: | Agarwal, Aarush, He, Raymond, Kieseler, Jan, Cremonesi, Matteo, Qasim, Shah Rukh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enabling Message Passing Interface Containers on the LUMI Supercomputer
by: Lazzaro, Alfio
Published: (2024)
by: Lazzaro, Alfio
Published: (2024)
pdGRASS: A Fast Parallel Density-Aware Algorithm for Graph Spectral Sparsification
by: Zhao, Tiancheng, et al.
Published: (2025)
by: Zhao, Tiancheng, et al.
Published: (2025)
GraphFlash: Enabling Fast and Elastic Graph Processing on Serverless Infrastructure
by: Zhao, Chen, et al.
Published: (2026)
by: Zhao, Chen, et al.
Published: (2026)
Fast Iterative Graph Computing with Updated Neighbor States
by: Zhou, Yijie, et al.
Published: (2024)
by: Zhou, Yijie, et al.
Published: (2024)
Efficient Graph Embedding at Scale: Optimizing CPU-GPU-SSD Integration
by: Li, Zhonggen, et al.
Published: (2025)
by: Li, Zhonggen, et al.
Published: (2025)
Towards Fast Setup and High Throughput of GPU Serverless Computing
by: Zhao, Han, et al.
Published: (2024)
by: Zhao, Han, et al.
Published: (2024)
Enhancing Cloud Task Scheduling Using a Hybrid Particle Swarm and Grey Wolf Optimization Approach
by: Prasad, Raveena, et al.
Published: (2025)
by: Prasad, Raveena, et al.
Published: (2025)
Byzantine Consensus in Directed Graphs with Message Authentication
by: Vaidya, Nitin H., et al.
Published: (2026)
by: Vaidya, Nitin H., et al.
Published: (2026)
cuFastTuckerPlus: A Stochastic Parallel Sparse FastTucker Decomposition Using GPU Tensor Cores
by: Li, Zixuan, et al.
Published: (2024)
by: Li, Zixuan, et al.
Published: (2024)
Parallel Track Transformers: Enabling Fast GPU Inference with Reduced Synchronization
by: Wang, Chong, et al.
Published: (2026)
by: Wang, Chong, et al.
Published: (2026)
Equivalence and Separation between Heard-Of and Asynchronous Message-Passing Models
by: Attiya, Hagit, et al.
Published: (2025)
by: Attiya, Hagit, et al.
Published: (2025)
HiRace: Accurate and Fast Source-Level Race Checking of GPU Programs
by: Jacobson, John, et al.
Published: (2024)
by: Jacobson, John, et al.
Published: (2024)
Fast Byzantine Total Order Broadcast
by: Monti, Matteo, et al.
Published: (2024)
by: Monti, Matteo, et al.
Published: (2024)
FastTrack: GPU-Accelerated Tracking for Visual SLAM
by: Khabiri, Kimia, et al.
Published: (2025)
by: Khabiri, Kimia, et al.
Published: (2025)
Formal Specification for Fast ACS: Low-Latency File-Based Ordered Message Delivery at Scale
by: Gupta, Sushant Kumar, et al.
Published: (2025)
by: Gupta, Sushant Kumar, et al.
Published: (2025)
Efficient Accelerated Graph Edit Distance Computation on GPU
by: Dabah, Adel, et al.
Published: (2026)
by: Dabah, Adel, et al.
Published: (2026)
TurboFFT: A High-Performance Fast Fourier Transform with Fault Tolerance on GPU
by: Wu, Shixun, et al.
Published: (2024)
by: Wu, Shixun, et al.
Published: (2024)
GRNND: A GPU-Parallel Relative NN-Descent Algorithm for Efficient Approximate Nearest Neighbor Graph Construction
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Accelerating Intra-Node GPU-to-GPU Communication Through Multi-Path Transfers with CUDA Graphs
by: Sojoodi, Amirhossein, et al.
Published: (2026)
by: Sojoodi, Amirhossein, et al.
Published: (2026)
Parallel GPU-Enabled Algorithms for SpGEMM on Arbitrary Semirings with Hybrid Communication
by: McFarland, Thomas, et al.
Published: (2025)
by: McFarland, Thomas, et al.
Published: (2025)
AGAThA: Fast and Efficient GPU Acceleration of Guided Sequence Alignment for Long Read Mapping
by: Park, Seongyeon, et al.
Published: (2024)
by: Park, Seongyeon, et al.
Published: (2024)
SOLANET: Distributed Neighbor Graph Construction on GPU-Accelerated Systems
by: Iwabuchi, Keita, et al.
Published: (2026)
by: Iwabuchi, Keita, et al.
Published: (2026)
A Survey of Distributed Graph Algorithms on Massive Graphs
by: Meng, Lingkai, et al.
Published: (2024)
by: Meng, Lingkai, et al.
Published: (2024)
Structures and Techniques for Streaming Dynamic Graph Processing on Decentralized Message-Driven Systems
by: Chandio, Bibrak Qamar, et al.
Published: (2024)
by: Chandio, Bibrak Qamar, et al.
Published: (2024)
Fast Algorithms for Scheduling Many-body Correlation Functions on Accelerators
by: Selvitopi, Oguz, et al.
Published: (2025)
by: Selvitopi, Oguz, et al.
Published: (2025)
A Multi-Objective Framework for Optimizing GPU-Enabled VM Placement in Cloud Data Centers with Multi-Instance GPU Technology
by: Siavashi, Ahmad, et al.
Published: (2025)
by: Siavashi, Ahmad, et al.
Published: (2025)
DAWN: Matrix Operation-Optimized Algorithm for Shortest Paths Problem on Unweighted Graphs
by: Feng, Yelai, et al.
Published: (2022)
by: Feng, Yelai, et al.
Published: (2022)
Multi-core & GPU-based Balanced Butterfly Counting in Signed Bipartite Graphs
by: Kiran, Mekala, et al.
Published: (2026)
by: Kiran, Mekala, et al.
Published: (2026)
City-Scale Visibility Graph Analysis via GPU-Accelerated HyperBall
by: Hodge, Alex, et al.
Published: (2026)
by: Hodge, Alex, et al.
Published: (2026)
Generating Dynamic Graph Algorithms for Multiple Backends for a Graph DSL
by: Behera, Nibedita, et al.
Published: (2025)
by: Behera, Nibedita, et al.
Published: (2025)
Optimizing the Variant Calling Pipeline Execution on Human Genomes Using GPU-Enabled Machines
by: Kumar, Ajay, et al.
Published: (2025)
by: Kumar, Ajay, et al.
Published: (2025)
Beeping Deterministic CONGEST Algorithms in Graphs
by: Garncarek, Pawel, et al.
Published: (2025)
by: Garncarek, Pawel, et al.
Published: (2025)
λScale: Enabling Fast Scaling for Serverless Large Language Model Inference
by: Yu, Minchen, et al.
Published: (2025)
by: Yu, Minchen, et al.
Published: (2025)
FlowWalker: A Memory-efficient and High-performance GPU-based Dynamic Graph Random Walk Framework
by: Mei, Junyi, et al.
Published: (2024)
by: Mei, Junyi, et al.
Published: (2024)
Fast Topology-Aware Lossy Data Compression with Full Preservation of Critical Points and Local Order
by: Fallin, Alex, et al.
Published: (2026)
by: Fallin, Alex, et al.
Published: (2026)
RIPPLE++: An Incremental Framework for Efficient GNN Inference on Evolving Graphs
by: Naman, Pranjal, et al.
Published: (2026)
by: Naman, Pranjal, et al.
Published: (2026)
Parameterized Task Graph Scheduling Algorithm for Comparing Algorithmic Components
by: Coleman, Jared, et al.
Published: (2024)
by: Coleman, Jared, et al.
Published: (2024)
Fast and Scalable Mixed Precision Euclidean Distance Calculations Using GPU Tensor Cores
by: Curless, Brian, et al.
Published: (2025)
by: Curless, Brian, et al.
Published: (2025)
A Preliminary Study on Accelerating Simulation Optimization with GPU Implementation
by: He, Jinghai, et al.
Published: (2024)
by: He, Jinghai, et al.
Published: (2024)
HC-SpMM: Accelerating Sparse Matrix-Matrix Multiplication for Graphs with Hybrid GPU Cores
by: Li, Zhonggen, et al.
Published: (2024)
by: Li, Zhonggen, et al.
Published: (2024)
Similar Items
-
Enabling Message Passing Interface Containers on the LUMI Supercomputer
by: Lazzaro, Alfio
Published: (2024) -
pdGRASS: A Fast Parallel Density-Aware Algorithm for Graph Spectral Sparsification
by: Zhao, Tiancheng, et al.
Published: (2025) -
GraphFlash: Enabling Fast and Elastic Graph Processing on Serverless Infrastructure
by: Zhao, Chen, et al.
Published: (2026) -
Fast Iterative Graph Computing with Updated Neighbor States
by: Zhou, Yijie, et al.
Published: (2024) -
Efficient Graph Embedding at Scale: Optimizing CPU-GPU-SSD Integration
by: Li, Zhonggen, et al.
Published: (2025)