Performance-Driven Optimization of Parallel Breadth-First Search
Fuente:
arXiv
Guardado en:
| Autores principales: | Bhaskar, Marati, Kanakagiri, Raghavendra |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
WaterWise: Co-optimizing Carbon- and Water-Footprint Toward Environmentally Sustainable Cloud Computing
por: Jiang, Yankai, et al.
Publicado: (2025)
por: Jiang, Yankai, et al.
Publicado: (2025)
ThirstyFLOPS: Water Footprint Modeling and Analysis Toward Sustainable HPC Systems
por: Jiang, Yankai, et al.
Publicado: (2025)
por: Jiang, Yankai, et al.
Publicado: (2025)
Fused Breadth-First Probabilistic Traversals on Distributed GPU Systems
por: Neff, Reece, et al.
Publicado: (2023)
por: Neff, Reece, et al.
Publicado: (2023)
Efficient Parallel Execution of Blockchain Transactions Leveraging Conflict Specifications
por: Anjana, Parwat Singh, et al.
Publicado: (2025)
por: Anjana, Parwat Singh, et al.
Publicado: (2025)
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
por: Kanakagiri, Raghavendra, et al.
Publicado: (2023)
por: Kanakagiri, Raghavendra, et al.
Publicado: (2023)
ForgetMeNot: Understanding and Modeling the Impact of Forever Chemicals Toward Sustainable Large-Scale Computing
por: Roy, Rohan Basu, et al.
Publicado: (2025)
por: Roy, Rohan Basu, et al.
Publicado: (2025)
A New Execution Model and Executor for Adaptively Optimizing the Performance of Parallel Algorithms Using HPX Runtime System
por: Mohammadiporshokooh, Karame, et al.
Publicado: (2025)
por: Mohammadiporshokooh, Karame, et al.
Publicado: (2025)
Parallelized Multi-Agent Bayesian Optimization in Lava
por: Snyder, Shay, et al.
Publicado: (2024)
por: Snyder, Shay, et al.
Publicado: (2024)
The Merit of Simple Policies: Buying Performance With Parallelism and System Architecture
por: Yildiz, Mert, et al.
Publicado: (2025)
por: Yildiz, Mert, et al.
Publicado: (2025)
High-Performance Parallelization of Dijkstra's Algorithm Using MPI and CUDA
por: Song, Boyang
Publicado: (2025)
por: Song, Boyang
Publicado: (2025)
PISA: An Adversarial Approach To Comparing Task Graph Scheduling Algorithms
por: Coleman, Jared, et al.
Publicado: (2024)
por: Coleman, Jared, et al.
Publicado: (2024)
Optimizing View Change for Byzantine Fault Tolerance in Parallel Consensus
por: Xie, Yifei, et al.
Publicado: (2026)
por: Xie, Yifei, et al.
Publicado: (2026)
Staleness-Centric Optimizations for Parallel Diffusion MoE Inference
por: Luo, Jiajun, et al.
Publicado: (2024)
por: Luo, Jiajun, et al.
Publicado: (2024)
Parallel DNA Sequence Alignment on High-Performance Systems with CUDA and MPI
por: Zwaka, Linus
Publicado: (2024)
por: Zwaka, Linus
Publicado: (2024)
Design Principles of Dynamic Resource Management for High-Performance Parallel Programming Models
por: Huber, Dominik, et al.
Publicado: (2024)
por: Huber, Dominik, et al.
Publicado: (2024)
Balancing Pipeline Parallelism with Vocabulary Parallelism
por: Yeung, Man Tsung, et al.
Publicado: (2024)
por: Yeung, Man Tsung, et al.
Publicado: (2024)
Heimdall++: Optimizing GPU Utilization and Pipeline Parallelism for Efficient Single-Pulse Detection
por: Xia, Bingzheng, et al.
Publicado: (2025)
por: Xia, Bingzheng, et al.
Publicado: (2025)
Optimizing Long-context LLM Serving via Fine-grained Sequence Parallelism
por: Li, Cong, et al.
Publicado: (2025)
por: Li, Cong, et al.
Publicado: (2025)
CFP: Efficient Optimization of Intra-Operator Parallelism Plans for Large Model Training
por: Hu, Weifang, et al.
Publicado: (2025)
por: Hu, Weifang, et al.
Publicado: (2025)
QPOPSS: Query and Parallelism Optimized Space-Saving for Finding Frequent Stream Elements
por: Jarlow, Victor, et al.
Publicado: (2024)
por: Jarlow, Victor, et al.
Publicado: (2024)
Committee Configuration Optimization for Parallel Byzantine Consensus in a Trusted Execution Environment
por: Xie, Yifei, et al.
Publicado: (2026)
por: Xie, Yifei, et al.
Publicado: (2026)
A Parallel CPU-GPU Framework for Batching Heuristic Operations in Depth-First Heuristic Search
por: Futuhi, Ehsan, et al.
Publicado: (2025)
por: Futuhi, Ehsan, et al.
Publicado: (2025)
Parallel Reduced Order Modeling for Digital Twins using High-Performance Computing Workflows
por: de Parga, S. Ares, et al.
Publicado: (2024)
por: de Parga, S. Ares, et al.
Publicado: (2024)
Automated Programmatic Performance Analysis of Parallel Programs
por: Cankur, Onur, et al.
Publicado: (2024)
por: Cankur, Onur, et al.
Publicado: (2024)
Parallel Data Object Creation: Towards Scalable Metadata Management in High-Performance I/O Library
por: Li, Youjia, et al.
Publicado: (2025)
por: Li, Youjia, et al.
Publicado: (2025)
STELLAR: Storage Tuning Engine Leveraging LLM Autonomous Reasoning for High Performance Parallel File Systems
por: Egersdoerfer, Chris, et al.
Publicado: (2026)
por: Egersdoerfer, Chris, et al.
Publicado: (2026)
A Comparative Review of Parallel Exact, Heuristic, Metaheuristic, and Hybrid Optimization Techniques for the Traveling Salesman Problem
por: Alkhalifa, Rabab, et al.
Publicado: (2025)
por: Alkhalifa, Rabab, et al.
Publicado: (2025)
The Entropy of Parallel Systems
por: Adefemi, Temitayo
Publicado: (2025)
por: Adefemi, Temitayo
Publicado: (2025)
Lectures on Parallel Computing
por: Träff, Jesper Larsson
Publicado: (2024)
por: Träff, Jesper Larsson
Publicado: (2024)
CaPGNN: Optimizing Parallel Graph Neural Network Training with Joint Caching and Resource-Aware Graph Partitioning
por: Song, Xianfeng, et al.
Publicado: (2025)
por: Song, Xianfeng, et al.
Publicado: (2025)
CoCoDiff: Optimizing Collective Communications for Distributed Diffusion Transformer Inference Under Ulysses Sequence Parallelism
por: Ma, Bin, et al.
Publicado: (2026)
por: Ma, Bin, et al.
Publicado: (2026)
NanoCP: Request-Level Dynamic Context Parallelism for Data-Expert Parallel Decoding
por: Chen, Jiefei, et al.
Publicado: (2026)
por: Chen, Jiefei, et al.
Publicado: (2026)
ZeroPP: Unleashing Exceptional Parallelism Efficiency through Tensor-Parallelism-Free Methodology
por: Tang, Ding, et al.
Publicado: (2024)
por: Tang, Ding, et al.
Publicado: (2024)
EvoSort: A Genetic-Algorithm-Based Adaptive Parallel Sorting Framework for Large-Scale High Performance Computing
por: Raj, Shashank, et al.
Publicado: (2025)
por: Raj, Shashank, et al.
Publicado: (2025)
Synergistic Tensor and Pipeline Parallelism
por: Qi, Mengshi, et al.
Publicado: (2025)
por: Qi, Mengshi, et al.
Publicado: (2025)
Distributed Locking: Performance Analysis and Optimization Strategies
por: Rodriguez, Andre, et al.
Publicado: (2025)
por: Rodriguez, Andre, et al.
Publicado: (2025)
Distributed Inference Performance Optimization for LLMs on CPUs
por: He, Pujiang, et al.
Publicado: (2024)
por: He, Pujiang, et al.
Publicado: (2024)
Looking for (Genomic) Needles in a Haystack: Sparsity-Driven Search for Identifying Correlated Genetic Mutations in Cancer
por: Prabhu, Ritvik, et al.
Publicado: (2026)
por: Prabhu, Ritvik, et al.
Publicado: (2026)
Parallel Seismic Data Processing Performance with Cloud-based Storage
por: Mohapatra, Sasmita, et al.
Publicado: (2025)
por: Mohapatra, Sasmita, et al.
Publicado: (2025)
Performance Optimization in Stream Processing Systems: Experiment-Driven Configuration Tuning for Kafka Streams
por: Chen, David, et al.
Publicado: (2026)
por: Chen, David, et al.
Publicado: (2026)
Ejemplares similares
-
WaterWise: Co-optimizing Carbon- and Water-Footprint Toward Environmentally Sustainable Cloud Computing
por: Jiang, Yankai, et al.
Publicado: (2025) -
ThirstyFLOPS: Water Footprint Modeling and Analysis Toward Sustainable HPC Systems
por: Jiang, Yankai, et al.
Publicado: (2025) -
Fused Breadth-First Probabilistic Traversals on Distributed GPU Systems
por: Neff, Reece, et al.
Publicado: (2023) -
Efficient Parallel Execution of Blockchain Transactions Leveraging Conflict Specifications
por: Anjana, Parwat Singh, et al.
Publicado: (2025) -
Minimum Cost Loop Nests for Contraction of a Sparse Tensor with a Tensor Network
por: Kanakagiri, Raghavendra, et al.
Publicado: (2023)