A Study on the Performance of Distributed Training of Data-driven CFD Simulations
Fuente:
arXiv
Saved in:
| Main Authors: | Iserte, Sergio, González-Barberá, Alejandro, Barreda, Paloma, Rojek, Krzysztof |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptation of AI-accelerated CFD Simulations to the IPU platform
by: Rosciszewski, P., et al.
Published: (2026)
by: Rosciszewski, P., et al.
Published: (2026)
Resource Optimization with MPI Process Malleability for Dynamic Workloads in HPC Clusters
by: Iserte, Sergio, et al.
Published: (2025)
by: Iserte, Sergio, et al.
Published: (2025)
Towards the Democratization and Standardization of Dynamic Resources with MPI Spawning
by: Iserte, Sergio, et al.
Published: (2026)
by: Iserte, Sergio, et al.
Published: (2026)
Malleable Molecular Dynamics Simulations with GROMACS and DMR
by: Sandås, Petter, et al.
Published: (2026)
by: Sandås, Petter, et al.
Published: (2026)
Wave-Based Dispatch for Circuit Cutting in Hybrid HPC--Quantum Systems
by: García-Raigada, Ricard S., et al.
Published: (2026)
by: García-Raigada, Ricard S., et al.
Published: (2026)
Leveraging Teaching on Demand: Approaching HPC to Undergrads
by: Catalán, S., et al.
Published: (2026)
by: Catalán, S., et al.
Published: (2026)
DMRlib: Easy-coding and Efficient Resource Management for Job Malleability
by: Iserte, Sergio, et al.
Published: (2026)
by: Iserte, Sergio, et al.
Published: (2026)
On the Performance and Memory Footprint of Distributed Training: An Empirical Study on Transformers
by: Lu, Zhengxian, et al.
Published: (2024)
by: Lu, Zhengxian, et al.
Published: (2024)
DFLOP: A Data-driven Framework for Multimodal LLM Training Pipeline Optimization
by: An, Hyeonjun, et al.
Published: (2026)
by: An, Hyeonjun, et al.
Published: (2026)
Empowering Distributed Training with Sparsity-driven Data Synchronization
by: Wang, Zhuang, et al.
Published: (2023)
by: Wang, Zhuang, et al.
Published: (2023)
MPI Malleability Validation under Replayed Real-World HPC Conditions
by: Iserte, S., et al.
Published: (2026)
by: Iserte, S., et al.
Published: (2026)
A Survey of End-to-End Modeling for Distributed DNN Training: Workloads, Simulators, and TCO
by: Svedas, Jonas, et al.
Published: (2025)
by: Svedas, Jonas, et al.
Published: (2025)
PALM: A Efficient Performance Simulator for Tiled Accelerators with Large-scale Model Training
by: Fang, Jiahao, et al.
Published: (2024)
by: Fang, Jiahao, et al.
Published: (2024)
Multi-Partner Project: Multi-GPU Performance Portability Analysis for CFD Simulations at Scale
by: Eleftherakis, Panagiotis-Eleftherios, et al.
Published: (2026)
by: Eleftherakis, Panagiotis-Eleftherios, et al.
Published: (2026)
A Test Taxonomy and Continuous Integration Ecosystem for Dynamic Resource Management in HPC
by: Sandås, Petter, et al.
Published: (2026)
by: Sandås, Petter, et al.
Published: (2026)
miniLB: A Performance Portability Study of Lattice-Boltzmann Simulations
by: Crisci, Luigi, et al.
Published: (2024)
by: Crisci, Luigi, et al.
Published: (2024)
PRISM: Probabilistic Runtime Insights and Scalable Performance Modeling for Large-Scale Distributed Training
by: Golden, Alicia, et al.
Published: (2025)
by: Golden, Alicia, et al.
Published: (2025)
DeFT: Mitigating Data Dependencies for Flexible Communication Scheduling in Distributed Training
by: Meng, Lin, et al.
Published: (2025)
by: Meng, Lin, et al.
Published: (2025)
An Explorative Study on Distributed Computing Techniques in Training and Inference of Large Language Models
by: Hakim, Sheikh Azizul, et al.
Published: (2025)
by: Hakim, Sheikh Azizul, et al.
Published: (2025)
A Framework for Consistency Models in Distributed Systems
by: Almeida, Paulo Sérgio
Published: (2024)
by: Almeida, Paulo Sérgio
Published: (2024)
Optimizing Data Distribution and Kernel Performance for Efficient Training of Chemistry Foundation Models: A Case Study with MACE
by: Firoz, Jesun, et al.
Published: (2025)
by: Firoz, Jesun, et al.
Published: (2025)
FUSCO: High-Performance Distributed Data Shuffling via Transformation-Communication Fusion
by: Zhu, Zhuoran, et al.
Published: (2025)
by: Zhu, Zhuoran, et al.
Published: (2025)
Ensuring Data Privacy in AC Optimal Power Flow with a Distributed Co-Simulation Framework
by: Dai, Xinliang, et al.
Published: (2024)
by: Dai, Xinliang, et al.
Published: (2024)
Enhancing Energy Efficiency in Scientific Workflows through CFD based PIVAEs
by: Zahir, Ali, et al.
Published: (2026)
by: Zahir, Ali, et al.
Published: (2026)
Distributed Computation with Local Advice
by: Balliu, Alkida, et al.
Published: (2024)
by: Balliu, Alkida, et al.
Published: (2024)
Cost-Performance Analysis: A Comparative Study of CPU-Based Serverless and GPU-Based Training Architectures
by: Barrak, Amine, et al.
Published: (2025)
by: Barrak, Amine, et al.
Published: (2025)
Echo: Simulating Distributed Training At Scale
by: Feng, Yicheng, et al.
Published: (2024)
by: Feng, Yicheng, et al.
Published: (2024)
Efficient Distributed MLLM Training with Cornstarch
by: Jang, Insu, et al.
Published: (2025)
by: Jang, Insu, et al.
Published: (2025)
DreamDDP: Accelerating Data Parallel Distributed LLM Training with Layer-wise Scheduled Partial Synchronization
by: Tang, Zhenheng, et al.
Published: (2025)
by: Tang, Zhenheng, et al.
Published: (2025)
A Comparative Analysis of Distributed Training Strategies for GPT-2
by: Patwardhan, Ishan, et al.
Published: (2024)
by: Patwardhan, Ishan, et al.
Published: (2024)
Automated Calibration of Parallel and Distributed Computing Simulators: A Case Study
by: McDonald, Jesse, et al.
Published: (2024)
by: McDonald, Jesse, et al.
Published: (2024)
A Distributed Edge FLISR Solution & Network Simulation Test Platform
by: Leniston, Darren, et al.
Published: (2024)
by: Leniston, Darren, et al.
Published: (2024)
Distributed Ranges: A Model for Distributed Data Structures, Algorithms, and Views
by: Brock, Benjamin, et al.
Published: (2024)
by: Brock, Benjamin, et al.
Published: (2024)
Efficient Training of Large Language Models on Distributed Infrastructures: A Survey
by: Duan, Jiangfei, et al.
Published: (2024)
by: Duan, Jiangfei, et al.
Published: (2024)
Distributing Context-Aware Shared Memory Data Structures: A Case Study on Singly-Linked Lists
by: Ravishankar, Raaghav, et al.
Published: (2024)
by: Ravishankar, Raaghav, et al.
Published: (2024)
Addressing Variable Heterogeneity in Distributed Multimodal Training with Entrain
by: Jang, Insu, et al.
Published: (2026)
by: Jang, Insu, et al.
Published: (2026)
Galvatron: Automatic Distributed Training for Large Transformer Models
by: Gumaan, Esmail
Published: (2025)
by: Gumaan, Esmail
Published: (2025)
Accelerating Distributed MoE Training and Inference with Lina
by: Li, Jiamin, et al.
Published: (2022)
by: Li, Jiamin, et al.
Published: (2022)
Optimizing Distributed Training Approaches for Scaling Neural Networks
by: Baligodugula, Vishnu Vardhan, et al.
Published: (2025)
by: Baligodugula, Vishnu Vardhan, et al.
Published: (2025)
Heta: Distributed Training of Heterogeneous Graph Neural Networks
by: Zhong, Yuchen, et al.
Published: (2024)
by: Zhong, Yuchen, et al.
Published: (2024)
Similar Items
-
Adaptation of AI-accelerated CFD Simulations to the IPU platform
by: Rosciszewski, P., et al.
Published: (2026) -
Resource Optimization with MPI Process Malleability for Dynamic Workloads in HPC Clusters
by: Iserte, Sergio, et al.
Published: (2025) -
Towards the Democratization and Standardization of Dynamic Resources with MPI Spawning
by: Iserte, Sergio, et al.
Published: (2026) -
Malleable Molecular Dynamics Simulations with GROMACS and DMR
by: Sandås, Petter, et al.
Published: (2026) -
Wave-Based Dispatch for Circuit Cutting in Hybrid HPC--Quantum Systems
by: García-Raigada, Ricard S., et al.
Published: (2026)