High-Performance Sorting-Based k-mer Counting in Distributed Memory with Flexible Hybrid Parallelism
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yifan, Guidi, Giulia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Asynchronous Distributed-Memory Parallel Algorithm for k-mer Counting
von: Hati, Souvadra, et al.
Veröffentlicht: (2025)
von: Hati, Souvadra, et al.
Veröffentlicht: (2025)
gpuPairHMM: High-speed Pair-HMM Forward Algorithm for DNA Variant Calling on GPUs
von: Schmidt, Bertil, et al.
Veröffentlicht: (2024)
von: Schmidt, Bertil, et al.
Veröffentlicht: (2024)
PanDelos-plus: A parallel algorithm for computing sequence homology in pangenomic analysis
von: Colli, Simone, et al.
Veröffentlicht: (2025)
von: Colli, Simone, et al.
Veröffentlicht: (2025)
MegIS: High-Performance, Energy-Efficient, and Low-Cost Metagenomic Analysis with In-Storage Processing
von: Ghiasi, Nika Mansouri, et al.
Veröffentlicht: (2024)
von: Ghiasi, Nika Mansouri, et al.
Veröffentlicht: (2024)
NMP-PaK: Near-Memory Processing Acceleration of Scalable De Novo Genome Assembly
von: Kim, Heewoo, et al.
Veröffentlicht: (2025)
von: Kim, Heewoo, et al.
Veröffentlicht: (2025)
RUBICON: A Framework for Designing Efficient Deep Learning-Based Genomic Basecallers
von: Singh, Gagandeep, et al.
Veröffentlicht: (2022)
von: Singh, Gagandeep, et al.
Veröffentlicht: (2022)
Parallel GPU-Enabled Algorithms for SpGEMM on Arbitrary Semirings with Hybrid Communication
von: McFarland, Thomas, et al.
Veröffentlicht: (2025)
von: McFarland, Thomas, et al.
Veröffentlicht: (2025)
Lock-free de Bruijn graph
von: Górniak, Daniel, et al.
Veröffentlicht: (2024)
von: Górniak, Daniel, et al.
Veröffentlicht: (2024)
TorchGWAS : GPU-accelerated GWAS for thousands of quantitative phenotypes
von: Zhao, Xingzhong, et al.
Veröffentlicht: (2026)
von: Zhao, Xingzhong, et al.
Veröffentlicht: (2026)
SAGe: A Lightweight Algorithm-Architecture Co-Design for Mitigating the Data Preparation Bottleneck in Large-Scale Genome Sequence Analysis
von: Ghiasi, Nika Mansouri, et al.
Veröffentlicht: (2025)
von: Ghiasi, Nika Mansouri, et al.
Veröffentlicht: (2025)
Ocean: Fast Estimation-Based Sparse General Matrix-Matrix Multiplication on GPU
von: Li, Yifan, et al.
Veröffentlicht: (2026)
von: Li, Yifan, et al.
Veröffentlicht: (2026)
EvoSort: A Genetic-Algorithm-Based Adaptive Parallel Sorting Framework for Large-Scale High Performance Computing
von: Raj, Shashank, et al.
Veröffentlicht: (2025)
von: Raj, Shashank, et al.
Veröffentlicht: (2025)
Efficient Chromosome Parallelization for Precision Medicine Genomic Workflows
von: Montserrat, Daniel Mas, et al.
Veröffentlicht: (2025)
von: Montserrat, Daniel Mas, et al.
Veröffentlicht: (2025)
Hybrid-Parallel: Achieving High Performance and Energy Efficient Distributed Inference on Robots
von: Sun, Zekai, et al.
Veröffentlicht: (2024)
von: Sun, Zekai, et al.
Veröffentlicht: (2024)
Performance Trade-offs of High Order Meshless Approximation on Distributed Memory Systems
von: Vehovar, Jon, et al.
Veröffentlicht: (2025)
von: Vehovar, Jon, et al.
Veröffentlicht: (2025)
On the Performance and Memory Footprint of Distributed Training: An Empirical Study on Transformers
von: Lu, Zhengxian, et al.
Veröffentlicht: (2024)
von: Lu, Zhengxian, et al.
Veröffentlicht: (2024)
Distributed-Memory Parallel Algorithms for Sparse Matrix and Sparse Tall-and-Skinny Matrix Multiplication
von: Ranawaka, Isuru, et al.
Veröffentlicht: (2024)
von: Ranawaka, Isuru, et al.
Veröffentlicht: (2024)
MixServe: An Automatic Distributed Serving System for MoE Models with Hybrid Parallelism Based on Fused Communication Algorithm
von: Zhou, Bowen, et al.
Veröffentlicht: (2026)
von: Zhou, Bowen, et al.
Veröffentlicht: (2026)
High-Performance Parallelization of Dijkstra's Algorithm Using MPI and CUDA
von: Song, Boyang
Veröffentlicht: (2025)
von: Song, Boyang
Veröffentlicht: (2025)
Justin: Hybrid CPU/Memory Elastic Scaling for Distributed Stream Processing
von: Schmitz, Donatien, et al.
Veröffentlicht: (2025)
von: Schmitz, Donatien, et al.
Veröffentlicht: (2025)
Cache Blocking of Distributed-Memory Parallel Matrix Power Kernels
von: Lacey, Dane C., et al.
Veröffentlicht: (2024)
von: Lacey, Dane C., et al.
Veröffentlicht: (2024)
A Flexible Programmable Pipeline Parallelism Framework for Efficient DNN Training
von: Jiang, Lijuan, et al.
Veröffentlicht: (2025)
von: Jiang, Lijuan, et al.
Veröffentlicht: (2025)
Parallel DNA Sequence Alignment on High-Performance Systems with CUDA and MPI
von: Zwaka, Linus
Veröffentlicht: (2024)
von: Zwaka, Linus
Veröffentlicht: (2024)
Accelerating Heterogeneous Tensor Parallelism via Flexible Workload Control
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
Design Principles of Dynamic Resource Management for High-Performance Parallel Programming Models
von: Huber, Dominik, et al.
Veröffentlicht: (2024)
von: Huber, Dominik, et al.
Veröffentlicht: (2024)
Parallel Data Object Creation: Towards Scalable Metadata Management in High-Performance I/O Library
von: Li, Youjia, et al.
Veröffentlicht: (2025)
von: Li, Youjia, et al.
Veröffentlicht: (2025)
Popcorn: Accelerating Kernel K-means on GPUs through Sparse Linear Algebra
von: Bellavita, Julian, et al.
Veröffentlicht: (2025)
von: Bellavita, Julian, et al.
Veröffentlicht: (2025)
HAP: Hybrid Adaptive Parallelism for Efficient Mixture-of-Experts Inference
von: Lin, Haoran, et al.
Veröffentlicht: (2025)
von: Lin, Haoran, et al.
Veröffentlicht: (2025)
Malleus: Straggler-Resilient Hybrid Parallel Training of Large-scale Models via Malleable Data and Model Parallelization
von: Li, Haoyang, et al.
Veröffentlicht: (2024)
von: Li, Haoyang, et al.
Veröffentlicht: (2024)
Parallel Reduced Order Modeling for Digital Twins using High-Performance Computing Workflows
von: de Parga, S. Ares, et al.
Veröffentlicht: (2024)
von: de Parga, S. Ares, et al.
Veröffentlicht: (2024)
Parallel Integer Sort: Theory and Practice
von: Dong, Xiaojun, et al.
Veröffentlicht: (2024)
von: Dong, Xiaojun, et al.
Veröffentlicht: (2024)
DHP: Efficient Scaling of MLLM Training with Dynamic Hybrid Parallelism
von: Niu, Yifan, et al.
Veröffentlicht: (2026)
von: Niu, Yifan, et al.
Veröffentlicht: (2026)
SiDP: Memory-Efficient Data Parallelism for Offline LLM Inference
von: Zhao, Alan, et al.
Veröffentlicht: (2026)
von: Zhao, Alan, et al.
Veröffentlicht: (2026)
DawnPiper: A Memory-scablable Pipeline Parallel Training Framework
von: Peng, Xuan, et al.
Veröffentlicht: (2025)
von: Peng, Xuan, et al.
Veröffentlicht: (2025)
DWDP: Distributed Weight Data Parallelism for High-Performance LLM Inference on NVL72
von: Li, Wanqian, et al.
Veröffentlicht: (2026)
von: Li, Wanqian, et al.
Veröffentlicht: (2026)
DynaFlow: Transparent and Flexible Intra-Device Parallelism via Programmable Operator Scheduling
von: Pan, Yi, et al.
Veröffentlicht: (2026)
von: Pan, Yi, et al.
Veröffentlicht: (2026)
Adaptra: Straggler-Resilient Hybrid-Parallel Training with Pipeline Adaptation
von: Wu, Tianyuan, et al.
Veröffentlicht: (2025)
von: Wu, Tianyuan, et al.
Veröffentlicht: (2025)
Experimental Evaluation of Distributed k-Core Decomposition
von: Guo, Bin, et al.
Veröffentlicht: (2024)
von: Guo, Bin, et al.
Veröffentlicht: (2024)
FUSCO: High-Performance Distributed Data Shuffling via Transformation-Communication Fusion
von: Zhu, Zhuoran, et al.
Veröffentlicht: (2025)
von: Zhu, Zhuoran, et al.
Veröffentlicht: (2025)
STELLAR: Storage Tuning Engine Leveraging LLM Autonomous Reasoning for High Performance Parallel File Systems
von: Egersdoerfer, Chris, et al.
Veröffentlicht: (2026)
von: Egersdoerfer, Chris, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
An Asynchronous Distributed-Memory Parallel Algorithm for k-mer Counting
von: Hati, Souvadra, et al.
Veröffentlicht: (2025) -
gpuPairHMM: High-speed Pair-HMM Forward Algorithm for DNA Variant Calling on GPUs
von: Schmidt, Bertil, et al.
Veröffentlicht: (2024) -
PanDelos-plus: A parallel algorithm for computing sequence homology in pangenomic analysis
von: Colli, Simone, et al.
Veröffentlicht: (2025) -
MegIS: High-Performance, Energy-Efficient, and Low-Cost Metagenomic Analysis with In-Storage Processing
von: Ghiasi, Nika Mansouri, et al.
Veröffentlicht: (2024) -
NMP-PaK: Near-Memory Processing Acceleration of Scalable De Novo Genome Assembly
von: Kim, Heewoo, et al.
Veröffentlicht: (2025)