Performance Trade-offs of High Order Meshless Approximation on Distributed Memory Systems
Fuente:
arXiv
Saved in:
| Main Authors: | Vehovar, Jon, Rot, Miha, Kosec, Gregor |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Load Balanced Parallel Node Generation for Meshless Numerical Methods
by: Vehovar, Jon, et al.
Published: (2026)
by: Vehovar, Jon, et al.
Published: (2026)
Space-Time Trade-off in Bounded Iterated Memory
by: Toyos-Marfurt, Guillermo, et al.
Published: (2025)
by: Toyos-Marfurt, Guillermo, et al.
Published: (2025)
Computation-Bandwidth-Memory Trade-offs: A Unified Paradigm for AI Infrastructure
by: Fan, Yuankai, et al.
Published: (2025)
by: Fan, Yuankai, et al.
Published: (2025)
CHIRON: Accelerating Node Synchronization without Security Trade-offs in Distributed Ledgers
by: Neiheiser, Ray, et al.
Published: (2024)
by: Neiheiser, Ray, et al.
Published: (2024)
On the Performance and Memory Footprint of Distributed Training: An Empirical Study on Transformers
by: Lu, Zhengxian, et al.
Published: (2024)
by: Lu, Zhengxian, et al.
Published: (2024)
Understanding Inference Scaling for LLMs: Bottlenecks, Trade-offs, and Performance Principles
by: Arif, Moiz, et al.
Published: (2026)
by: Arif, Moiz, et al.
Published: (2026)
System-Level Performance Modeling of Photonic In-Memory Computing
by: Arockiaraj, Jebacyril, et al.
Published: (2026)
by: Arockiaraj, Jebacyril, et al.
Published: (2026)
Exploring Performance-Productivity Trade-offs in AMT Runtimes: A Task Bench Study of Itoyori, ItoyoriFBC, HPX, and MPI
by: Lahnor, Torben R., et al.
Published: (2026)
by: Lahnor, Torben R., et al.
Published: (2026)
Self-Evolving Distributed Memory Architecture for Scalable AI Systems
by: Li, Zixuan, et al.
Published: (2026)
by: Li, Zixuan, et al.
Published: (2026)
A Study on Messaging Trade-offs in Data Streaming for Scientific Workflows
by: George, Anjus, et al.
Published: (2025)
by: George, Anjus, et al.
Published: (2025)
High-Performance Sorting-Based k-mer Counting in Distributed Memory with Flexible Hybrid Parallelism
by: Li, Yifan, et al.
Published: (2024)
by: Li, Yifan, et al.
Published: (2024)
Parallel Reduced Order Modeling for Digital Twins using High-Performance Computing Workflows
by: de Parga, S. Ares, et al.
Published: (2024)
by: de Parga, S. Ares, et al.
Published: (2024)
Approximate Byzantine Fault-Tolerance in Distributed Optimization
by: Liu, Shuo, et al.
Published: (2021)
by: Liu, Shuo, et al.
Published: (2021)
The Complexity of Distributed Minimum Weight Cycle Approximation
by: Chang, Yi-Jun, et al.
Published: (2026)
by: Chang, Yi-Jun, et al.
Published: (2026)
A Space-Time Trade-off for Fast Self-Stabilizing Leader Election in Population Protocols
by: Austin, Henry, et al.
Published: (2025)
by: Austin, Henry, et al.
Published: (2025)
To Compress or Not To Compress: Energy Trade-Offs and Benefits of Lossy Compressed I/O
by: Wilkins, Grant, et al.
Published: (2024)
by: Wilkins, Grant, et al.
Published: (2024)
NestedFP: High-Performance, Memory-Efficient Dual-Precision Floating Point Support for LLMs
by: Lee, Haeun, et al.
Published: (2025)
by: Lee, Haeun, et al.
Published: (2025)
FUSCO: High-Performance Distributed Data Shuffling via Transformation-Communication Fusion
by: Zhu, Zhuoran, et al.
Published: (2025)
by: Zhu, Zhuoran, et al.
Published: (2025)
Preparing for HPC on RISC-V: Examining Vectorization and Distributed Performance of an Astrophyiscs Application with HPX and Kokkos
by: Diehl, Patrick, et al.
Published: (2024)
by: Diehl, Patrick, et al.
Published: (2024)
Solutions for Distributed Memory Access Mechanism on HPC Clusters
by: Meizner, Jan, et al.
Published: (2025)
by: Meizner, Jan, et al.
Published: (2025)
Optimizing Memory Allocation in Distributed Clusters with Predictive Modeling
by: Bader, Jonathan, et al.
Published: (2026)
by: Bader, Jonathan, et al.
Published: (2026)
madupite: A High-Performance Distributed Solver for Large-Scale Markov Decision Processes
by: Gargiani, Matilde, et al.
Published: (2025)
by: Gargiani, Matilde, et al.
Published: (2025)
Performance of Distributed File Systems on Cloud Computing Environment: An Evaluation for Small-File Problem
by: Duong, Thanh, et al.
Published: (2023)
by: Duong, Thanh, et al.
Published: (2023)
Shifting the Sweet Spot: High-Performance Matrix-Free Method for High-Order Elasticity
by: Chang, Dali, et al.
Published: (2026)
by: Chang, Dali, et al.
Published: (2026)
PathWeaver: A High-Throughput Multi-GPU System for Graph-Based Approximate Nearest Neighbor Search
by: Kim, Sukjin, et al.
Published: (2025)
by: Kim, Sukjin, et al.
Published: (2025)
Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows
by: Ifath, Md. Monzurul Amin, et al.
Published: (2026)
by: Ifath, Md. Monzurul Amin, et al.
Published: (2026)
CXL Shared Memory Programming: Barely Distributed and Almost Persistent
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
Efficient and Portable Support for Overdecomposition on Distributed Memory GPGPU Platforms
by: Bhosale, Aditya, et al.
Published: (2026)
by: Bhosale, Aditya, et al.
Published: (2026)
Parallel DNA Sequence Alignment on High-Performance Systems with CUDA and MPI
by: Zwaka, Linus
Published: (2024)
by: Zwaka, Linus
Published: (2024)
Edge AI in Highly Volatile Environments: Is Fairness Worth the Accuracy Trade-off?
by: Zaland, Obaidullah, et al.
Published: (2025)
by: Zaland, Obaidullah, et al.
Published: (2025)
Approximated Coded Computing: Towards Fast, Private and Secure Distributed Machine Learning
by: Qiu, Houming, et al.
Published: (2024)
by: Qiu, Houming, et al.
Published: (2024)
Justin: Hybrid CPU/Memory Elastic Scaling for Distributed Stream Processing
by: Schmitz, Donatien, et al.
Published: (2025)
by: Schmitz, Donatien, et al.
Published: (2025)
PULSE: Accelerating Distributed Pointer-Traversals on Disaggregated Memory (Extended Version)
by: Tang, Yupeng, et al.
Published: (2023)
by: Tang, Yupeng, et al.
Published: (2023)
Distributed Order Recording Techniques for Efficient Record-and-Replay of Multi-threaded Programs
by: Fu, Xiang, et al.
Published: (2026)
by: Fu, Xiang, et al.
Published: (2026)
Learning-Augmented Performance Model for Tensor Product Factorization in High-Order FEM
by: Ren, Xuanzhengbo, et al.
Published: (2026)
by: Ren, Xuanzhengbo, et al.
Published: (2026)
Predictive Performance of Photonic SRAM-based In-Memory Computing for Tensor Decomposition
by: Wijeratne, Sasindu, et al.
Published: (2025)
by: Wijeratne, Sasindu, et al.
Published: (2025)
Shared Virtual Memory: Its Design and Performance Implications for Diverse Applications
by: Cooper, Bennett, et al.
Published: (2024)
by: Cooper, Bennett, et al.
Published: (2024)
The Design and Implementation of a High-Performance Log-Structured RAID System for ZNS SSDs
by: Li, Jinhong, et al.
Published: (2024)
by: Li, Jinhong, et al.
Published: (2024)
Distributed Locking: Performance Analysis and Optimization Strategies
by: Rodriguez, Andre, et al.
Published: (2025)
by: Rodriguez, Andre, et al.
Published: (2025)
Distributed Inference Performance Optimization for LLMs on CPUs
by: He, Pujiang, et al.
Published: (2024)
by: He, Pujiang, et al.
Published: (2024)
Similar Items
-
Load Balanced Parallel Node Generation for Meshless Numerical Methods
by: Vehovar, Jon, et al.
Published: (2026) -
Space-Time Trade-off in Bounded Iterated Memory
by: Toyos-Marfurt, Guillermo, et al.
Published: (2025) -
Computation-Bandwidth-Memory Trade-offs: A Unified Paradigm for AI Infrastructure
by: Fan, Yuankai, et al.
Published: (2025) -
CHIRON: Accelerating Node Synchronization without Security Trade-offs in Distributed Ledgers
by: Neiheiser, Ray, et al.
Published: (2024) -
On the Performance and Memory Footprint of Distributed Training: An Empirical Study on Transformers
by: Lu, Zhengxian, et al.
Published: (2024)