Performance-Driven Optimization of Parallel Breadth-First Search
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhaskar, Marati, Kanakagiri, Raghavendra |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WaterWise: Co-optimizing Carbon- and Water-Footprint Toward Environmentally Sustainable Cloud Computing
von: Jiang, Yankai, et al.
Veröffentlicht: (2025)
von: Jiang, Yankai, et al.
Veröffentlicht: (2025)
ThirstyFLOPS: Water Footprint Modeling and Analysis Toward Sustainable HPC Systems
von: Jiang, Yankai, et al.
Veröffentlicht: (2025)
von: Jiang, Yankai, et al.
Veröffentlicht: (2025)
Fused Breadth-First Probabilistic Traversals on Distributed GPU Systems
von: Neff, Reece, et al.
Veröffentlicht: (2023)
von: Neff, Reece, et al.
Veröffentlicht: (2023)
Efficient Parallel Execution of Blockchain Transactions Leveraging Conflict Specifications
von: Anjana, Parwat Singh, et al.
Veröffentlicht: (2025)
von: Anjana, Parwat Singh, et al.
Veröffentlicht: (2025)
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
von: Kanakagiri, Raghavendra, et al.
Veröffentlicht: (2023)
von: Kanakagiri, Raghavendra, et al.
Veröffentlicht: (2023)
ForgetMeNot: Understanding and Modeling the Impact of Forever Chemicals Toward Sustainable Large-Scale Computing
von: Roy, Rohan Basu, et al.
Veröffentlicht: (2025)
von: Roy, Rohan Basu, et al.
Veröffentlicht: (2025)
A New Execution Model and Executor for Adaptively Optimizing the Performance of Parallel Algorithms Using HPX Runtime System
von: Mohammadiporshokooh, Karame, et al.
Veröffentlicht: (2025)
von: Mohammadiporshokooh, Karame, et al.
Veröffentlicht: (2025)
Parallelized Multi-Agent Bayesian Optimization in Lava
von: Snyder, Shay, et al.
Veröffentlicht: (2024)
von: Snyder, Shay, et al.
Veröffentlicht: (2024)
The Merit of Simple Policies: Buying Performance With Parallelism and System Architecture
von: Yildiz, Mert, et al.
Veröffentlicht: (2025)
von: Yildiz, Mert, et al.
Veröffentlicht: (2025)
High-Performance Parallelization of Dijkstra's Algorithm Using MPI and CUDA
von: Song, Boyang
Veröffentlicht: (2025)
von: Song, Boyang
Veröffentlicht: (2025)
PISA: An Adversarial Approach To Comparing Task Graph Scheduling Algorithms
von: Coleman, Jared, et al.
Veröffentlicht: (2024)
von: Coleman, Jared, et al.
Veröffentlicht: (2024)
Optimizing View Change for Byzantine Fault Tolerance in Parallel Consensus
von: Xie, Yifei, et al.
Veröffentlicht: (2026)
von: Xie, Yifei, et al.
Veröffentlicht: (2026)
Staleness-Centric Optimizations for Parallel Diffusion MoE Inference
von: Luo, Jiajun, et al.
Veröffentlicht: (2024)
von: Luo, Jiajun, et al.
Veröffentlicht: (2024)
Parallel DNA Sequence Alignment on High-Performance Systems with CUDA and MPI
von: Zwaka, Linus
Veröffentlicht: (2024)
von: Zwaka, Linus
Veröffentlicht: (2024)
Design Principles of Dynamic Resource Management for High-Performance Parallel Programming Models
von: Huber, Dominik, et al.
Veröffentlicht: (2024)
von: Huber, Dominik, et al.
Veröffentlicht: (2024)
Balancing Pipeline Parallelism with Vocabulary Parallelism
von: Yeung, Man Tsung, et al.
Veröffentlicht: (2024)
von: Yeung, Man Tsung, et al.
Veröffentlicht: (2024)
Heimdall++: Optimizing GPU Utilization and Pipeline Parallelism for Efficient Single-Pulse Detection
von: Xia, Bingzheng, et al.
Veröffentlicht: (2025)
von: Xia, Bingzheng, et al.
Veröffentlicht: (2025)
Optimizing Long-context LLM Serving via Fine-grained Sequence Parallelism
von: Li, Cong, et al.
Veröffentlicht: (2025)
von: Li, Cong, et al.
Veröffentlicht: (2025)
CFP: Efficient Optimization of Intra-Operator Parallelism Plans for Large Model Training
von: Hu, Weifang, et al.
Veröffentlicht: (2025)
von: Hu, Weifang, et al.
Veröffentlicht: (2025)
QPOPSS: Query and Parallelism Optimized Space-Saving for Finding Frequent Stream Elements
von: Jarlow, Victor, et al.
Veröffentlicht: (2024)
von: Jarlow, Victor, et al.
Veröffentlicht: (2024)
Committee Configuration Optimization for Parallel Byzantine Consensus in a Trusted Execution Environment
von: Xie, Yifei, et al.
Veröffentlicht: (2026)
von: Xie, Yifei, et al.
Veröffentlicht: (2026)
A Parallel CPU-GPU Framework for Batching Heuristic Operations in Depth-First Heuristic Search
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2025)
von: Futuhi, Ehsan, et al.
Veröffentlicht: (2025)
Parallel Reduced Order Modeling for Digital Twins using High-Performance Computing Workflows
von: de Parga, S. Ares, et al.
Veröffentlicht: (2024)
von: de Parga, S. Ares, et al.
Veröffentlicht: (2024)
Automated Programmatic Performance Analysis of Parallel Programs
von: Cankur, Onur, et al.
Veröffentlicht: (2024)
von: Cankur, Onur, et al.
Veröffentlicht: (2024)
Parallel Data Object Creation: Towards Scalable Metadata Management in High-Performance I/O Library
von: Li, Youjia, et al.
Veröffentlicht: (2025)
von: Li, Youjia, et al.
Veröffentlicht: (2025)
STELLAR: Storage Tuning Engine Leveraging LLM Autonomous Reasoning for High Performance Parallel File Systems
von: Egersdoerfer, Chris, et al.
Veröffentlicht: (2026)
von: Egersdoerfer, Chris, et al.
Veröffentlicht: (2026)
A Comparative Review of Parallel Exact, Heuristic, Metaheuristic, and Hybrid Optimization Techniques for the Traveling Salesman Problem
von: Alkhalifa, Rabab, et al.
Veröffentlicht: (2025)
von: Alkhalifa, Rabab, et al.
Veröffentlicht: (2025)
The Entropy of Parallel Systems
von: Adefemi, Temitayo
Veröffentlicht: (2025)
von: Adefemi, Temitayo
Veröffentlicht: (2025)
Lectures on Parallel Computing
von: Träff, Jesper Larsson
Veröffentlicht: (2024)
von: Träff, Jesper Larsson
Veröffentlicht: (2024)
CaPGNN: Optimizing Parallel Graph Neural Network Training with Joint Caching and Resource-Aware Graph Partitioning
von: Song, Xianfeng, et al.
Veröffentlicht: (2025)
von: Song, Xianfeng, et al.
Veröffentlicht: (2025)
CoCoDiff: Optimizing Collective Communications for Distributed Diffusion Transformer Inference Under Ulysses Sequence Parallelism
von: Ma, Bin, et al.
Veröffentlicht: (2026)
von: Ma, Bin, et al.
Veröffentlicht: (2026)
NanoCP: Request-Level Dynamic Context Parallelism for Data-Expert Parallel Decoding
von: Chen, Jiefei, et al.
Veröffentlicht: (2026)
von: Chen, Jiefei, et al.
Veröffentlicht: (2026)
ZeroPP: Unleashing Exceptional Parallelism Efficiency through Tensor-Parallelism-Free Methodology
von: Tang, Ding, et al.
Veröffentlicht: (2024)
von: Tang, Ding, et al.
Veröffentlicht: (2024)
EvoSort: A Genetic-Algorithm-Based Adaptive Parallel Sorting Framework for Large-Scale High Performance Computing
von: Raj, Shashank, et al.
Veröffentlicht: (2025)
von: Raj, Shashank, et al.
Veröffentlicht: (2025)
Synergistic Tensor and Pipeline Parallelism
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
von: Qi, Mengshi, et al.
Veröffentlicht: (2025)
Distributed Locking: Performance Analysis and Optimization Strategies
von: Rodriguez, Andre, et al.
Veröffentlicht: (2025)
von: Rodriguez, Andre, et al.
Veröffentlicht: (2025)
Distributed Inference Performance Optimization for LLMs on CPUs
von: He, Pujiang, et al.
Veröffentlicht: (2024)
von: He, Pujiang, et al.
Veröffentlicht: (2024)
Looking for (Genomic) Needles in a Haystack: Sparsity-Driven Search for Identifying Correlated Genetic Mutations in Cancer
von: Prabhu, Ritvik, et al.
Veröffentlicht: (2026)
von: Prabhu, Ritvik, et al.
Veröffentlicht: (2026)
Parallel Seismic Data Processing Performance with Cloud-based Storage
von: Mohapatra, Sasmita, et al.
Veröffentlicht: (2025)
von: Mohapatra, Sasmita, et al.
Veröffentlicht: (2025)
Performance Optimization in Stream Processing Systems: Experiment-Driven Configuration Tuning for Kafka Streams
von: Chen, David, et al.
Veröffentlicht: (2026)
von: Chen, David, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
WaterWise: Co-optimizing Carbon- and Water-Footprint Toward Environmentally Sustainable Cloud Computing
von: Jiang, Yankai, et al.
Veröffentlicht: (2025) -
ThirstyFLOPS: Water Footprint Modeling and Analysis Toward Sustainable HPC Systems
von: Jiang, Yankai, et al.
Veröffentlicht: (2025) -
Fused Breadth-First Probabilistic Traversals on Distributed GPU Systems
von: Neff, Reece, et al.
Veröffentlicht: (2023) -
Efficient Parallel Execution of Blockchain Transactions Leveraging Conflict Specifications
von: Anjana, Parwat Singh, et al.
Veröffentlicht: (2025) -
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
von: Kanakagiri, Raghavendra, et al.
Veröffentlicht: (2023)