Beyond Exascale: Dataflow Domain Translation on a Cerebras Cluster
Fuente:
arXiv
Saved in:
| Main Authors: | Oppelstrup, Tomas, Giamblanco, Nicholas, Kalchev, Delyan Z., Sharapov, Ilya, Taylor, Mark, Van Essendelft, Dirk, Rajamanickam, Sivasankaran, James, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Breaking the Molecular Dynamics Timescale Barrier Using a Wafer-Scale System
by: Santos, Kylee, et al.
Published: (2024)
by: Santos, Kylee, et al.
Published: (2024)
Breaking the mold: overcoming the time constraints of molecular dynamics on general-purpose hardware
by: Perez, Danny, et al.
Published: (2024)
by: Perez, Danny, et al.
Published: (2024)
LAPIS: A Performance Portable, High Productivity Compiler Framework
by: Kelley, Brian, et al.
Published: (2025)
by: Kelley, Brian, et al.
Published: (2025)
CELLO: Co-designing Schedule and Hybrid Implicit/Explicit Buffer for Complex Tensor Reuse
by: Garg, Raveesh, et al.
Published: (2023)
by: Garg, Raveesh, et al.
Published: (2023)
Jet: Multilevel Graph Partitioning on Graphics Processing Units
by: Gilbert, Michael S., et al.
Published: (2023)
by: Gilbert, Michael S., et al.
Published: (2023)
Exploring Sparse Matrix Multiplication Kernels on the Cerebras CS-3
by: Shah, Milan, et al.
Published: (2026)
by: Shah, Milan, et al.
Published: (2026)
Benchmarking the Performance of Large Language Models on the Cerebras Wafer Scale Engine
by: Zhang, Zuoning, et al.
Published: (2024)
by: Zhang, Zuoning, et al.
Published: (2024)
A System Level Compiler for Massively-Parallel, Spatial, Dataflow Architectures
by: Van Essendelft, Dirk, et al.
Published: (2025)
by: Van Essendelft, Dirk, et al.
Published: (2025)
Towards Exascale Computation for Turbomachinery Flows
by: Fu, Yuhang, et al.
Published: (2023)
by: Fu, Yuhang, et al.
Published: (2023)
Enhancing Real-Time Master Data Management with Complex Match and Merge Algorithms
by: Rajamanickam, Durai
Published: (2024)
by: Rajamanickam, Durai
Published: (2024)
Stencil Computations on Cerebras Wafer-Scale Engine
by: Belli, Elia, et al.
Published: (2026)
by: Belli, Elia, et al.
Published: (2026)
Asynchronous-Many-Task Systems: Challenges and Opportunities -- Scaling an AMR Astrophysics Code on Exascale machines using Kokkos and HPX
by: Daiß, Gregor, et al.
Published: (2024)
by: Daiß, Gregor, et al.
Published: (2024)
Exploration of Energy and Throughput Tradeoffs for Dataflow Networks
by: Karim, Abrarul, et al.
Published: (2026)
by: Karim, Abrarul, et al.
Published: (2026)
Towards Exascale Computing for Astrophysical Simulation Leveraging the Leonardo EuroHPC System
by: Shukla, Nitin, et al.
Published: (2025)
by: Shukla, Nitin, et al.
Published: (2025)
Sustaining Exascale Performance: Lessons from HPL and HPL-MxP on Aurora
by: Goto, Kazushige, et al.
Published: (2026)
by: Goto, Kazushige, et al.
Published: (2026)
Mapping Gemma3 onto an Edge Dataflow Architecture
by: Du, Shouyu, et al.
Published: (2026)
by: Du, Shouyu, et al.
Published: (2026)
The Renoir Dataflow Platform: Efficient Data Processing without Complexity
by: De Martini, Luca, et al.
Published: (2023)
by: De Martini, Luca, et al.
Published: (2023)
Dataflow-Oriented Classification and Performance Analysis of GPU-Accelerated Homomorphic Encryption
by: Nozaki, Ai, et al.
Published: (2026)
by: Nozaki, Ai, et al.
Published: (2026)
SUNDIALS Time Integrators for Exascale Applications with Many Independent ODE Systems
by: Balos, Cody J., et al.
Published: (2024)
by: Balos, Cody J., et al.
Published: (2024)
Enhancing Performance Insight at Scale: A Heterogeneous Framework for Exascale Diagnostics
by: Grbic, Dragana
Published: (2026)
by: Grbic, Dragana
Published: (2026)
Styx: Transactional Stateful Functions on Streaming Dataflows
by: Psarakis, Kyriakos, et al.
Published: (2023)
by: Psarakis, Kyriakos, et al.
Published: (2023)
DEEP: Edge-based Dataflow Processing with Hybrid Docker Hub and Regional Registries
by: Mehran, Narges, et al.
Published: (2025)
by: Mehran, Narges, et al.
Published: (2025)
TileLoom: Automatic Dataflow Planning for Tile-Based Languages on Spatial Dataflow Accelerators
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
Employing Artificial Intelligence to Steer Exascale Workflows with Colmena
by: Ward, Logan, et al.
Published: (2024)
by: Ward, Logan, et al.
Published: (2024)
CheckMate: Evaluating Checkpointing Protocols for Streaming Dataflows
by: Siachamis, George, et al.
Published: (2024)
by: Siachamis, George, et al.
Published: (2024)
Exascale In-situ visualization for Astronomy & Cosmology
by: Tuccari, Nicola, et al.
Published: (2025)
by: Tuccari, Nicola, et al.
Published: (2025)
FLARE: A Dataflow-Aware and Scalable Hardware Architecture for Neural-Hybrid Scientific Lossy Compression
by: Jia, Wenqi, et al.
Published: (2025)
by: Jia, Wenqi, et al.
Published: (2025)
Stateful Entities: Object-oriented Cloud Applications as Distributed Dataflows
by: Psarakis, Kyriakos, et al.
Published: (2021)
by: Psarakis, Kyriakos, et al.
Published: (2021)
Kitsune: Enabling Dataflow Execution on GPUs
by: Davies, Michael, et al.
Published: (2025)
by: Davies, Michael, et al.
Published: (2025)
Suki: Choreographed Distributed Dataflow in Rust
by: Laddad, Shadaj, et al.
Published: (2024)
by: Laddad, Shadaj, et al.
Published: (2024)
Democratizing Scalable Cloud Applications: Transactional Stateful Functions on Streaming Dataflows
by: Psarakis, Kyriakos
Published: (2025)
by: Psarakis, Kyriakos
Published: (2025)
A Digital Twin Framework for Liquid-cooled Supercomputers as Demonstrated at Exascale
by: Brewer, Wesley, et al.
Published: (2024)
by: Brewer, Wesley, et al.
Published: (2024)
CMDS: Cross-layer Dataflow Optimization for DNN Accelerators Exploiting Multi-bank Memories
by: Shi, Man, et al.
Published: (2024)
by: Shi, Man, et al.
Published: (2024)
Alya towards Exascale: Optimal OpenACC Performance of the Navier-Stokes Finite Element Assembly on GPUs
by: Owen, Herbert, et al.
Published: (2024)
by: Owen, Herbert, et al.
Published: (2024)
Experiences Porting Distributed Applications to Asynchronous Tasks: A Multidimensional FFT Case-study
by: Strack, Alexander, et al.
Published: (2024)
by: Strack, Alexander, et al.
Published: (2024)
DGNNFlow: A Streaming Dataflow Architecture for Real-Time Edge-based Dynamic GNN Inference in HL-LHC Trigger Systems
by: Maharaj, Davendra, et al.
Published: (2026)
by: Maharaj, Davendra, et al.
Published: (2026)
Failure Transparency in Stateful Dataflow Systems (Technical Report)
by: Veresov, Aleksey, et al.
Published: (2024)
by: Veresov, Aleksey, et al.
Published: (2024)
Beyond A Single AI Cluster: A Survey of Decentralized LLM Training
by: Dong, Haotian, et al.
Published: (2025)
by: Dong, Haotian, et al.
Published: (2025)
Fine-Grained Power and Energy Attribution on AMD GPU/APU-Based Exascale Nodes
by: McDaniel, Adam, et al.
Published: (2026)
by: McDaniel, Adam, et al.
Published: (2026)
FlowUnits: Extending Dataflow for the Edge-to-Cloud Computing Continuum
by: Chini, Fabio, et al.
Published: (2025)
by: Chini, Fabio, et al.
Published: (2025)
Similar Items
-
Breaking the Molecular Dynamics Timescale Barrier Using a Wafer-Scale System
by: Santos, Kylee, et al.
Published: (2024) -
Breaking the mold: overcoming the time constraints of molecular dynamics on general-purpose hardware
by: Perez, Danny, et al.
Published: (2024) -
LAPIS: A Performance Portable, High Productivity Compiler Framework
by: Kelley, Brian, et al.
Published: (2025) -
CELLO: Co-designing Schedule and Hybrid Implicit/Explicit Buffer for Complex Tensor Reuse
by: Garg, Raveesh, et al.
Published: (2023) -
Jet: Multilevel Graph Partitioning on Graphics Processing Units
by: Gilbert, Michael S., et al.
Published: (2023)