ML-based Modeling to Predict I/O Performance on Different Storage Sub-systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Yiheng, Sivaraman, Pranav, Devarajan, Hariharan, Mohror, Kathryn, Bhatele, Abhinav |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Recorder: Comprehensive Parallel I/O Tracing and Analysis
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Taking GPU Programming Models to Task for Performance Portability
di: Davis, Joshua H., et al.
Pubblicazione: (2024)
di: Davis, Joshua H., et al.
Pubblicazione: (2024)
Formal Definitions and Performance Comparison of Consistency Models for Parallel File Systems
di: Wang, Chen, et al.
Pubblicazione: (2024)
di: Wang, Chen, et al.
Pubblicazione: (2024)
Analytics of Longitudinal System Monitoring Data for Performance Prediction
di: Costello, Ian J., et al.
Pubblicazione: (2020)
di: Costello, Ian J., et al.
Pubblicazione: (2020)
Characterizing Production GPU Workloads using System-wide Telemetry Data
di: Cankur, Onur, et al.
Pubblicazione: (2025)
di: Cankur, Onur, et al.
Pubblicazione: (2025)
ParEval-Repo: A Benchmark Suite for Evaluating LLMs with Repository-level HPC Translation Tasks
di: Davis, Joshua H., et al.
Pubblicazione: (2025)
di: Davis, Joshua H., et al.
Pubblicazione: (2025)
Performance-Aligned LLMs for Generating Fast Code
di: Nichols, Daniel, et al.
Pubblicazione: (2024)
di: Nichols, Daniel, et al.
Pubblicazione: (2024)
Integrating Performance Tools in Model Reasoning for GPU Kernel Optimization
di: Nichols, Daniel, et al.
Pubblicazione: (2025)
di: Nichols, Daniel, et al.
Pubblicazione: (2025)
Automated Programmatic Performance Analysis of Parallel Programs
di: Cankur, Onur, et al.
Pubblicazione: (2024)
di: Cankur, Onur, et al.
Pubblicazione: (2024)
HPC-Coder: Modeling Parallel Programs using Large Language Models
di: Nichols, Daniel, et al.
Pubblicazione: (2023)
di: Nichols, Daniel, et al.
Pubblicazione: (2023)
Pipit: Scripting the analysis of parallel execution traces
di: Bhatele, Abhinav, et al.
Pubblicazione: (2023)
di: Bhatele, Abhinav, et al.
Pubblicazione: (2023)
KEET: Explaining Performance of GPU Kernels Using LLM Agents
di: Davis, Joshua H., et al.
Pubblicazione: (2026)
di: Davis, Joshua H., et al.
Pubblicazione: (2026)
Can Large Language Models Write Parallel Code?
di: Nichols, Daniel, et al.
Pubblicazione: (2024)
di: Nichols, Daniel, et al.
Pubblicazione: (2024)
The Big Send-off: Scalable and Performant Collectives for Deep Learning
di: Singh, Siddharth, et al.
Pubblicazione: (2025)
di: Singh, Siddharth, et al.
Pubblicazione: (2025)
Understanding and Improving Communication Performance in Multi-node LLM Inference
di: Singhania, Prajwal, et al.
Pubblicazione: (2025)
di: Singhania, Prajwal, et al.
Pubblicazione: (2025)
Performance Models for a Two-tiered Storage System
di: Sasidharan, Aparna, et al.
Pubblicazione: (2025)
di: Sasidharan, Aparna, et al.
Pubblicazione: (2025)
HPC-Coder-V2: Studying Code LLMs Across Low-Resource Parallel Languages
di: Chaturvedi, Aman, et al.
Pubblicazione: (2024)
di: Chaturvedi, Aman, et al.
Pubblicazione: (2024)
Pandemics In Silico: Scaling an Agent-Based Simulation on Realistic Social Contact Networks
di: Kitson, Joy, et al.
Pubblicazione: (2024)
di: Kitson, Joy, et al.
Pubblicazione: (2024)
Evaluating Cross-Architecture Performance Modeling of Distributed ML Workloads Using StableHLO
di: Svedas, Jonas, et al.
Pubblicazione: (2026)
di: Svedas, Jonas, et al.
Pubblicazione: (2026)
Plexus: Taming Billion-edge Graphs with 3D Parallel Full-graph GNN Training
di: Ranjan, Aditya K., et al.
Pubblicazione: (2025)
di: Ranjan, Aditya K., et al.
Pubblicazione: (2025)
HPAC-ML: A Programming Model for Embedding ML Surrogates in Scientific Applications
di: Fink, Zane, et al.
Pubblicazione: (2024)
di: Fink, Zane, et al.
Pubblicazione: (2024)
Coded Data Rebalancing for Distributed Data Storage Systems with Cyclic Storage
di: Vaishya, Abhinav, et al.
Pubblicazione: (2022)
di: Vaishya, Abhinav, et al.
Pubblicazione: (2022)
A Performance Analyzer for a Public Cloud's ML-Augmented VM Allocator
di: Bostandoost, Roozbeh, et al.
Pubblicazione: (2025)
di: Bostandoost, Roozbeh, et al.
Pubblicazione: (2025)
IOAgent: Democratizing Trustworthy HPC I/O Performance Diagnosis Capability via LLMs
di: Egersdoerfer, Chris, et al.
Pubblicazione: (2026)
di: Egersdoerfer, Chris, et al.
Pubblicazione: (2026)
Parallel Seismic Data Processing Performance with Cloud-based Storage
di: Mohapatra, Sasmita, et al.
Pubblicazione: (2025)
di: Mohapatra, Sasmita, et al.
Pubblicazione: (2025)
Use Cases for High Performance Research Desktops
di: Henschel, Robert, et al.
Pubblicazione: (2024)
di: Henschel, Robert, et al.
Pubblicazione: (2024)
ML-based Adaptive Prefetching and Data Placement for US HEP Systems
di: Karanam, Venkat Sai Suman Lamba, et al.
Pubblicazione: (2025)
di: Karanam, Venkat Sai Suman Lamba, et al.
Pubblicazione: (2025)
Predictive Performance of Photonic SRAM-based In-Memory Computing for Tensor Decomposition
di: Wijeratne, Sasindu, et al.
Pubblicazione: (2025)
di: Wijeratne, Sasindu, et al.
Pubblicazione: (2025)
Optimizing the Longhorn Cloud-native Software Defined Storage Engine for High Performance
di: Kampadais, Konstantinos, et al.
Pubblicazione: (2025)
di: Kampadais, Konstantinos, et al.
Pubblicazione: (2025)
Fast Kronecker Matrix-Matrix Multiplication on GPUs
di: Jangda, Abhinav, et al.
Pubblicazione: (2024)
di: Jangda, Abhinav, et al.
Pubblicazione: (2024)
CubicML: Automated ML for Large ML Systems Co-design with ML Prediction of Performance
di: Wen, Wei, et al.
Pubblicazione: (2024)
di: Wen, Wei, et al.
Pubblicazione: (2024)
PUSHtap: PIM-based In-Memory HTAP with Unified Data Storage Format
di: Zhao, Yilong, et al.
Pubblicazione: (2025)
di: Zhao, Yilong, et al.
Pubblicazione: (2025)
Adaptive Cache Management for Complex Storage Systems Using CNN-LSTM-Based Spatiotemporal Prediction
di: Wang, Xiaoye, et al.
Pubblicazione: (2024)
di: Wang, Xiaoye, et al.
Pubblicazione: (2024)
A Bring-Your-Own-Model Approach for ML-Driven Storage Placement in Warehouse-Scale Computers
di: Yang, Chenxi, et al.
Pubblicazione: (2025)
di: Yang, Chenxi, et al.
Pubblicazione: (2025)
Revisiting Speculative Leaderless Protocols for Low-Latency BFT Replication
di: Qian, Daniel, et al.
Pubblicazione: (2026)
di: Qian, Daniel, et al.
Pubblicazione: (2026)
STELLAR: Storage Tuning Engine Leveraging LLM Autonomous Reasoning for High Performance Parallel File Systems
di: Egersdoerfer, Chris, et al.
Pubblicazione: (2026)
di: Egersdoerfer, Chris, et al.
Pubblicazione: (2026)
Reducing the Impact of I/O Contention in Numerical Weather Prediction Workflows at Scale Using DAOS
di: Manubens, Nicolau, et al.
Pubblicazione: (2024)
di: Manubens, Nicolau, et al.
Pubblicazione: (2024)
Oblivious Robots Performing Different Tasks on Grid Without Knowing their Team Members
di: Ghosh, Satakshi, et al.
Pubblicazione: (2022)
di: Ghosh, Satakshi, et al.
Pubblicazione: (2022)
MSF-Model: Queuing-Based Analysis and Prediction of Metastable Failures in Replicated Storage Systems
di: Habibi, Farzad, et al.
Pubblicazione: (2023)
di: Habibi, Farzad, et al.
Pubblicazione: (2023)
Parallel Data Object Creation: Towards Scalable Metadata Management in High-Performance I/O Library
di: Li, Youjia, et al.
Pubblicazione: (2025)
di: Li, Youjia, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Recorder: Comprehensive Parallel I/O Tracing and Analysis
di: Wang, Chen, et al.
Pubblicazione: (2025) -
Taking GPU Programming Models to Task for Performance Portability
di: Davis, Joshua H., et al.
Pubblicazione: (2024) -
Formal Definitions and Performance Comparison of Consistency Models for Parallel File Systems
di: Wang, Chen, et al.
Pubblicazione: (2024) -
Analytics of Longitudinal System Monitoring Data for Performance Prediction
di: Costello, Ian J., et al.
Pubblicazione: (2020) -
Characterizing Production GPU Workloads using System-wide Telemetry Data
di: Cankur, Onur, et al.
Pubblicazione: (2025)