Saved in:
| Main Authors: | Park, Junseok, Maury, Eduardo A., Oh, Changhoon, Shin, Donghoon, Denisko, Danielle, Lee, Eunjung Alice |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.15377 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Parallel Downloader for Large Genomic Datasets
by: Swargo, Rasman Mubtasim, et al.
Published: (2025)
by: Swargo, Rasman Mubtasim, et al.
Published: (2025)
Optimizing the Variant Calling Pipeline Execution on Human Genomes Using GPU-Enabled Machines
by: Kumar, Ajay, et al.
Published: (2025)
by: Kumar, Ajay, et al.
Published: (2025)
Looking for (Genomic) Needles in a Haystack: Sparsity-Driven Search for Identifying Correlated Genetic Mutations in Cancer
by: Prabhu, Ritvik, et al.
Published: (2026)
by: Prabhu, Ritvik, et al.
Published: (2026)
Characterizing Compute-Communication Overlap in GPU-Accelerated Distributed Deep Learning: Performance and Power Implications
by: Lee, Seonho, et al.
Published: (2025)
by: Lee, Seonho, et al.
Published: (2025)
NAVIS: Concurrent Search and Update with Low Position-Seeking Overhead in On-SSD Graph-Based Vector Search
by: Song, Jaeyong, et al.
Published: (2026)
by: Song, Jaeyong, et al.
Published: (2026)
FlexiWalker: Extensible GPU Framework for Efficient Dynamic Random Walks with Runtime Adaptation
by: Park, Seongyeon, et al.
Published: (2025)
by: Park, Seongyeon, et al.
Published: (2025)
Breaking the Capacity Bottleneck in Model-Heterogeneous Federated Learning via Gradual Model Restoration
by: Ma, Chengjie, et al.
Published: (2025)
by: Ma, Chengjie, et al.
Published: (2025)
NestedFP: High-Performance, Memory-Efficient Dual-Precision Floating Point Support for LLMs
by: Lee, Haeun, et al.
Published: (2025)
by: Lee, Haeun, et al.
Published: (2025)
Eliminating Hidden Serialization in Multi-Node Megakernel Communication
by: Oh, Byungsoo, et al.
Published: (2026)
by: Oh, Byungsoo, et al.
Published: (2026)
Action Deviation-Aware Inference for Low-Latency Wireless Robots
by: Park, Jeyoung, et al.
Published: (2025)
by: Park, Jeyoung, et al.
Published: (2025)
Bringing computation to the data: A MOEA-driven approach for optimising data processing in the context of the SKA and SRCNet
by: Parra-Royón, Manuel, et al.
Published: (2026)
by: Parra-Royón, Manuel, et al.
Published: (2026)
WANify: Gauging and Balancing Runtime WAN Bandwidth for Geo-distributed Data Analytics
by: Mohapatra, Anshuman Das, et al.
Published: (2025)
by: Mohapatra, Anshuman Das, et al.
Published: (2025)
RUBICON: A Framework for Designing Efficient Deep Learning-Based Genomic Basecallers
by: Singh, Gagandeep, et al.
Published: (2022)
by: Singh, Gagandeep, et al.
Published: (2022)
Accelerating LLM Inference with Precomputed Query Storage
by: Park, Jay H., et al.
Published: (2025)
by: Park, Jay H., et al.
Published: (2025)
MPI-over-CXL: Enhancing Communication Efficiency in Distributed HPC Systems
by: Kwon, Miryeong, et al.
Published: (2025)
by: Kwon, Miryeong, et al.
Published: (2025)
SAGe: A Lightweight Algorithm-Architecture Co-Design for Mitigating the Data Preparation Bottleneck in Large-Scale Genome Sequence Analysis
by: Ghiasi, Nika Mansouri, et al.
Published: (2025)
by: Ghiasi, Nika Mansouri, et al.
Published: (2025)
NMP-PaK: Near-Memory Processing Acceleration of Scalable De Novo Genome Assembly
by: Kim, Heewoo, et al.
Published: (2025)
by: Kim, Heewoo, et al.
Published: (2025)
Litmus: Fair Pricing for Serverless Computing
by: Pei, Qi, et al.
Published: (2024)
by: Pei, Qi, et al.
Published: (2024)
AGAThA: Fast and Efficient GPU Acceleration of Guided Sequence Alignment for Long Read Mapping
by: Park, Seongyeon, et al.
Published: (2024)
by: Park, Seongyeon, et al.
Published: (2024)
Towards Exascale Computation for Turbomachinery Flows
by: Fu, Yuhang, et al.
Published: (2023)
by: Fu, Yuhang, et al.
Published: (2023)
Quantifying Autoscaler Vulnerabilities: An Empirical Study of Resource Misallocation Induced by Cloud Infrastructure Faults
by: Park, Gijun
Published: (2026)
by: Park, Gijun
Published: (2026)
PIM-SHERPA: Software Method for On-device LLM Inference by Resolving PIM Memory Attribute and Layout Inconsistencies
by: Lee, Sunjung, et al.
Published: (2026)
by: Lee, Sunjung, et al.
Published: (2026)
bittide: Control Time, Not Flows
by: Bastiaan, Martijn, et al.
Published: (2025)
by: Bastiaan, Martijn, et al.
Published: (2025)
CaGR-RAG: Context-aware Query Grouping for Disk-based Vector Search in RAG Systems
by: Jeong, Yeonwoo, et al.
Published: (2025)
by: Jeong, Yeonwoo, et al.
Published: (2025)
Application of cloud computing platform in industrial big data processing
by: Yao, Ziyan
Published: (2024)
by: Yao, Ziyan
Published: (2024)
Parallel Gaussian process with kernel approximation in CUDA
by: Carminati, Davide
Published: (2024)
by: Carminati, Davide
Published: (2024)
GriNNder: Breaking the Memory Capacity Wall in Full-Graph GNN Training with Storage Offloading
by: Song, Jaeyong, et al.
Published: (2026)
by: Song, Jaeyong, et al.
Published: (2026)
Accelerating Optimal Power Flow with GPUs: SIMD Abstraction of Nonlinear Programs and Condensed-Space Interior-Point Methods
by: Shin, Sungho, et al.
Published: (2023)
by: Shin, Sungho, et al.
Published: (2023)
Performance Modeling and Evaluation of Hyperledger Fabric: An Analysis Based on Transaction Flow and Endorsement Policies
by: Melo, Carlos, et al.
Published: (2025)
by: Melo, Carlos, et al.
Published: (2025)
MatKV: Trading Compute for Flash Storage in LLM Inference
by: Shin, Kun-Woo, et al.
Published: (2025)
by: Shin, Kun-Woo, et al.
Published: (2025)
PathWeaver: A High-Throughput Multi-GPU System for Graph-Based Approximate Nearest Neighbor Search
by: Kim, Sukjin, et al.
Published: (2025)
by: Kim, Sukjin, et al.
Published: (2025)
Majorum: Ebb-and-Flow Consensus with Dynamic Quorums
by: D'Amato, Francesco, et al.
Published: (2026)
by: D'Amato, Francesco, et al.
Published: (2026)
A Hierarchical Sharded Blockchain Balancing Performance and Availability
by: Jo, Yongrae, et al.
Published: (2025)
by: Jo, Yongrae, et al.
Published: (2025)
Efficient Long Context Fine-tuning with Chunk Flow
by: Yuan, Xiulong, et al.
Published: (2025)
by: Yuan, Xiulong, et al.
Published: (2025)
Scalable quality control on processing of large diffusion-weighted and structural magnetic resonance imaging datasets
by: Kim, Michael E., et al.
Published: (2024)
by: Kim, Michael E., et al.
Published: (2024)
Profiling and Modeling of Power Characteristics of Leadership-Scale HPC System Workloads
by: Karimi, Ahmad Maroof, et al.
Published: (2024)
by: Karimi, Ahmad Maroof, et al.
Published: (2024)
FlowMoE: A Scalable Pipeline Scheduling Framework for Distributed Mixture-of-Experts Training
by: Gao, Yunqi, et al.
Published: (2025)
by: Gao, Yunqi, et al.
Published: (2025)
FlowMesh: A Service Fabric for Composable LLM Workflows
by: Shen, Junyi, et al.
Published: (2025)
by: Shen, Junyi, et al.
Published: (2025)
Flow-Bench: A Dataset for Computational Workflow Anomaly Detection
by: Papadimitriou, George, et al.
Published: (2023)
by: Papadimitriou, George, et al.
Published: (2023)
Heterogeneous Federated Learning with Prototype Alignment and Upscaling
by: Lee, Gyuejeong, et al.
Published: (2025)
by: Lee, Gyuejeong, et al.
Published: (2025)
Similar Items
-
Adaptive Parallel Downloader for Large Genomic Datasets
by: Swargo, Rasman Mubtasim, et al.
Published: (2025) -
Optimizing the Variant Calling Pipeline Execution on Human Genomes Using GPU-Enabled Machines
by: Kumar, Ajay, et al.
Published: (2025) -
Looking for (Genomic) Needles in a Haystack: Sparsity-Driven Search for Identifying Correlated Genetic Mutations in Cancer
by: Prabhu, Ritvik, et al.
Published: (2026) -
Characterizing Compute-Communication Overlap in GPU-Accelerated Distributed Deep Learning: Performance and Power Implications
by: Lee, Seonho, et al.
Published: (2025) -
NAVIS: Concurrent Search and Update with Low Position-Seeking Overhead in On-SSD Graph-Based Vector Search
by: Song, Jaeyong, et al.
Published: (2026)