TorchGWAS : GPU-accelerated GWAS for thousands of quantitative phenotypes
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Xingzhong, Xie, Ziqian, Islam, Saiful, Sheikh Muhammad, Xia, Tian, Chen, Cheng, Zhi, Degui |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
gpuPairHMM: High-speed Pair-HMM Forward Algorithm for DNA Variant Calling on GPUs
by: Schmidt, Bertil, et al.
Published: (2024)
by: Schmidt, Bertil, et al.
Published: (2024)
An Asynchronous Distributed-Memory Parallel Algorithm for k-mer Counting
by: Hati, Souvadra, et al.
Published: (2025)
by: Hati, Souvadra, et al.
Published: (2025)
High-Performance Sorting-Based k-mer Counting in Distributed Memory with Flexible Hybrid Parallelism
by: Li, Yifan, et al.
Published: (2024)
by: Li, Yifan, et al.
Published: (2024)
PanDelos-plus: A parallel algorithm for computing sequence homology in pangenomic analysis
by: Colli, Simone, et al.
Published: (2025)
by: Colli, Simone, et al.
Published: (2025)
Efficiently Reproducing Distributed Workflows in Notebook-based Systems
by: Azaz, Talha, et al.
Published: (2026)
by: Azaz, Talha, et al.
Published: (2026)
MegIS: High-Performance, Energy-Efficient, and Low-Cost Metagenomic Analysis with In-Storage Processing
by: Ghiasi, Nika Mansouri, et al.
Published: (2024)
by: Ghiasi, Nika Mansouri, et al.
Published: (2024)
SAGe: A Lightweight Algorithm-Architecture Co-Design for Mitigating the Data Preparation Bottleneck in Large-Scale Genome Sequence Analysis
by: Ghiasi, Nika Mansouri, et al.
Published: (2025)
by: Ghiasi, Nika Mansouri, et al.
Published: (2025)
NMP-PaK: Near-Memory Processing Acceleration of Scalable De Novo Genome Assembly
by: Kim, Heewoo, et al.
Published: (2025)
by: Kim, Heewoo, et al.
Published: (2025)
Lock-free de Bruijn graph
by: Górniak, Daniel, et al.
Published: (2024)
by: Górniak, Daniel, et al.
Published: (2024)
RUBICON: A Framework for Designing Efficient Deep Learning-Based Genomic Basecallers
by: Singh, Gagandeep, et al.
Published: (2022)
by: Singh, Gagandeep, et al.
Published: (2022)
Pipelined Dense Symmetric Eigenvalue Decomposition on Multi-GPU Architectures
by: Wang, Hansheng, et al.
Published: (2025)
by: Wang, Hansheng, et al.
Published: (2025)
Toward Portable GPU Performance: Julia Recursive Implementation of TRMM and TRSM
by: Carrica, Vicki, et al.
Published: (2025)
by: Carrica, Vicki, et al.
Published: (2025)
Ocean: Fast Estimation-Based Sparse General Matrix-Matrix Multiplication on GPU
by: Li, Yifan, et al.
Published: (2026)
by: Li, Yifan, et al.
Published: (2026)
Implementing Multi-GPU Scientific Computing Miniapps Across Performance Portable Frameworks
by: Villalobos, Johansell, et al.
Published: (2025)
by: Villalobos, Johansell, et al.
Published: (2025)
Communication-Avoiding SpGEMM via Trident Partitioning on Hierarchical GPU Interconnects
by: Bellavita, Julian, et al.
Published: (2026)
by: Bellavita, Julian, et al.
Published: (2026)
Integrating Odeint Time Stepping into OpenFPM for Distributed and GPU Accelerated Numerical Solvers
by: Singh, Abhinav, et al.
Published: (2023)
by: Singh, Abhinav, et al.
Published: (2023)
Performant Unified GPU Kernels for Portable Singular Value Computation Across Hardware and Precision
by: Ringoot, Evelyne, et al.
Published: (2025)
by: Ringoot, Evelyne, et al.
Published: (2025)
Investigating Matrix Repartitioning to Address the Over- and Undersubscription Challenge for a GPU-based CFD Solver
by: Olenik, Gregor, et al.
Published: (2025)
by: Olenik, Gregor, et al.
Published: (2025)
Anomaly Detection in Large-Scale Cloud Systems: An Industry Case and Dataset
by: Islam, Mohammad Saiful, et al.
Published: (2024)
by: Islam, Mohammad Saiful, et al.
Published: (2024)
On the energy efficiency of sparse matrix computations on multi-GPU clusters
by: Bernaschi, Massimo, et al.
Published: (2025)
by: Bernaschi, Massimo, et al.
Published: (2025)
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
by: Nichols, Daniel, et al.
Published: (2025)
by: Nichols, Daniel, et al.
Published: (2025)
GPU Implementations for Midsize Integer Addition and Multiplication
by: Oancea, Cosmin E., et al.
Published: (2024)
by: Oancea, Cosmin E., et al.
Published: (2024)
Multi-Objective Load Balancing for Heterogeneous Edge-Based Object Detection Systems
by: Alqahtani, Daghash K., et al.
Published: (2026)
by: Alqahtani, Daghash K., et al.
Published: (2026)
Efficient Chromosome Parallelization for Precision Medicine Genomic Workflows
by: Montserrat, Daniel Mas, et al.
Published: (2025)
by: Montserrat, Daniel Mas, et al.
Published: (2025)
SPUMA: a minimally invasive approach to the GPU porting of OPENFOAM
by: Bnà, Simone, et al.
Published: (2025)
by: Bnà, Simone, et al.
Published: (2025)
Combining GPU and CPU for accelerating evolutionary computing workloads
by: Eynaliyev, Rustam, et al.
Published: (2025)
by: Eynaliyev, Rustam, et al.
Published: (2025)
GoldbachGPU: An Open Source GPU-Accelerated Framework for Verification of Goldbach's Conjecture
by: Llorente-Saguer, Isaac
Published: (2026)
by: Llorente-Saguer, Isaac
Published: (2026)
SGPRS: Seamless GPU Partitioning Real-Time Scheduler for Periodic Deep Learning Workloads
by: Babaei, Amir Fakhim, et al.
Published: (2024)
by: Babaei, Amir Fakhim, et al.
Published: (2024)
A GPU-accelerated Molecular Docking Workflow with Kubernetes and Apache Airflow
by: Medeiros, Daniel, et al.
Published: (2024)
by: Medeiros, Daniel, et al.
Published: (2024)
Supercharging Federated Learning with Flower and NVIDIA FLARE
by: Roth, Holger R., et al.
Published: (2024)
by: Roth, Holger R., et al.
Published: (2024)
A Study of Performance Programming of CPU, GPU accelerated Computers and SIMD Architecture
by: Yi, Xinyao
Published: (2024)
by: Yi, Xinyao
Published: (2024)
PETSc/TAO Developments for GPU-Based Early Exascale Systems
by: Mills, Richard Tran, et al.
Published: (2024)
by: Mills, Richard Tran, et al.
Published: (2024)
WgPy: GPU-accelerated NumPy-like array library for web browsers
by: Hidaka, Masatoshi, et al.
Published: (2025)
by: Hidaka, Masatoshi, et al.
Published: (2025)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
by: Diehl, Patrick, et al.
Published: (2025)
by: Diehl, Patrick, et al.
Published: (2025)
SeBS-Flow: Benchmarking Serverless Cloud Function Workflows
by: Schmid, Larissa, et al.
Published: (2024)
by: Schmid, Larissa, et al.
Published: (2024)
A Unifying Framework to Enable Artificial Intelligence in High Performance Computing Workflows
by: Domke, Jens, et al.
Published: (2025)
by: Domke, Jens, et al.
Published: (2025)
CloudHeatMap: Heatmap-Based Monitoring for Large-Scale Cloud Systems
by: Sohana, Sarah, et al.
Published: (2024)
by: Sohana, Sarah, et al.
Published: (2024)
$μ$OpTime: Statically Reducing the Execution Time of Microbenchmark Suites Using Stability Metrics
by: Japke, Nils, et al.
Published: (2025)
by: Japke, Nils, et al.
Published: (2025)
Adaptable TeaStore
by: Bliudze, Simon, et al.
Published: (2024)
by: Bliudze, Simon, et al.
Published: (2024)
A Test Taxonomy and Continuous Integration Ecosystem for Dynamic Resource Management in HPC
by: Sandås, Petter, et al.
Published: (2026)
by: Sandås, Petter, et al.
Published: (2026)
Similar Items
-
gpuPairHMM: High-speed Pair-HMM Forward Algorithm for DNA Variant Calling on GPUs
by: Schmidt, Bertil, et al.
Published: (2024) -
An Asynchronous Distributed-Memory Parallel Algorithm for k-mer Counting
by: Hati, Souvadra, et al.
Published: (2025) -
High-Performance Sorting-Based k-mer Counting in Distributed Memory with Flexible Hybrid Parallelism
by: Li, Yifan, et al.
Published: (2024) -
PanDelos-plus: A parallel algorithm for computing sequence homology in pangenomic analysis
by: Colli, Simone, et al.
Published: (2025) -
Efficiently Reproducing Distributed Workflows in Notebook-based Systems
by: Azaz, Talha, et al.
Published: (2026)