Mat2Boundary: Treating User-Defined Boundary Condition as SpMV for Distributed PDE Solvers on Block-Structured Grids
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cai, Yanzheng, Zhang, Mingzhe, Chen, Shengqi, Song, Haoyuan, Chen, Wenguang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CB-SpMV:A Data Aggregating and Balance Algorithm for Cache-Friendly Block-Based SpMV on GPUs
von: Cong, Xing, et al.
Veröffentlicht: (2026)
von: Cong, Xing, et al.
Veröffentlicht: (2026)
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage
von: Lin, Junqing, et al.
Veröffentlicht: (2025)
von: Lin, Junqing, et al.
Veröffentlicht: (2025)
MERBIT: A GPU-Based SpMV Method for Iterative Workloads
von: Zhang, Qi, et al.
Veröffentlicht: (2026)
von: Zhang, Qi, et al.
Veröffentlicht: (2026)
A Nonlinear Hash-based Optimization Method for SpMV on GPUs
von: Yan, Chen, et al.
Veröffentlicht: (2025)
von: Yan, Chen, et al.
Veröffentlicht: (2025)
FlashMP: Fast Discrete Transform-Based Solver for Preconditioning Maxwell's Equations on GPUs
von: Zhang, Haoyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyuan, et al.
Veröffentlicht: (2025)
SDSL-Solver: Scalable Distributed Sparse Linear Solvers for Large-Scale Interior Point Methods
von: Yang, Shaofeng, et al.
Veröffentlicht: (2026)
von: Yang, Shaofeng, et al.
Veröffentlicht: (2026)
FEPLB: Exploiting Copy Engines for Nearly Free MoE Load Balancing in Distributed Training
von: Qi, Shuyao, et al.
Veröffentlicht: (2026)
von: Qi, Shuyao, et al.
Veröffentlicht: (2026)
AES-SpMM: Balancing Accuracy and Speed by Adaptive Edge Sampling Strategy to Accelerate SpMM in GNNs
von: Song, Yingchen, et al.
Veröffentlicht: (2025)
von: Song, Yingchen, et al.
Veröffentlicht: (2025)
RSH-SpMM: A Row-Structured Hybrid Kernel for Sparse Matrix-Matrix Multiplication on GPUs
von: Li, Aiying, et al.
Veröffentlicht: (2026)
von: Li, Aiying, et al.
Veröffentlicht: (2026)
Sustainable Grid through Distributed Data Centers: Spinning AI Demand for Grid Stabilization and Optimization
von: Evans, Scott C, et al.
Veröffentlicht: (2025)
von: Evans, Scott C, et al.
Veröffentlicht: (2025)
PackSELL: A Sparse Matrix Format for Precision-Agnostic High-Performance SpMV
von: Suzuki, Kengo, et al.
Veröffentlicht: (2026)
von: Suzuki, Kengo, et al.
Veröffentlicht: (2026)
madupite: A High-Performance Distributed Solver for Large-Scale Markov Decision Processes
von: Gargiani, Matilde, et al.
Veröffentlicht: (2025)
von: Gargiani, Matilde, et al.
Veröffentlicht: (2025)
High Performance Unstructured SpMM Computation Using Tensor Cores
von: Okanovic, Patrik, et al.
Veröffentlicht: (2024)
von: Okanovic, Patrik, et al.
Veröffentlicht: (2024)
Improving SpGEMM Performance Through Matrix Reordering and Cluster-wise Computation
von: Islam, Abdullah Al Raqibul, et al.
Veröffentlicht: (2025)
von: Islam, Abdullah Al Raqibul, et al.
Veröffentlicht: (2025)
Parallel GPU-Enabled Algorithms for SpGEMM on Arbitrary Semirings with Hybrid Communication
von: McFarland, Thomas, et al.
Veröffentlicht: (2025)
von: McFarland, Thomas, et al.
Veröffentlicht: (2025)
Algebraic Temporal Blocking for Sparse Iterative Solvers on Multi-Core CPUs
von: Alappat, Christie, et al.
Veröffentlicht: (2023)
von: Alappat, Christie, et al.
Veröffentlicht: (2023)
ParamSpMM: Adaptive and Efficient Sparse Matrix-Matrix Multiplication on GPUs for GNNs
von: Zhang, Lixing, et al.
Veröffentlicht: (2026)
von: Zhang, Lixing, et al.
Veröffentlicht: (2026)
A Structure-Aware Irregular Blocking Method for Sparse LU Factorization
von: Hu, Zhen, et al.
Veröffentlicht: (2025)
von: Hu, Zhen, et al.
Veröffentlicht: (2025)
Distributed Variational Quantum Linear Solver
von: Lu, Chao, et al.
Veröffentlicht: (2026)
von: Lu, Chao, et al.
Veröffentlicht: (2026)
BlockRaFT: A Distributed Framework for Fault-Tolerant and Scalable Blockchain Nodes
von: Piduguralla, Manaswini, et al.
Veröffentlicht: (2026)
von: Piduguralla, Manaswini, et al.
Veröffentlicht: (2026)
AB-Sparse: Sparse Attention with Adaptive Block Size for Accurate and Efficient Long-Context Inference
von: Liu, Di, et al.
Veröffentlicht: (2026)
von: Liu, Di, et al.
Veröffentlicht: (2026)
HC-SpMM: Accelerating Sparse Matrix-Matrix Multiplication for Graphs with Hybrid GPU Cores
von: Li, Zhonggen, et al.
Veröffentlicht: (2024)
von: Li, Zhonggen, et al.
Veröffentlicht: (2024)
StableShard: Stable and Scalable Blockchain Sharding with High Concurrency via Collaborative Committees
von: Li, Mingzhe, et al.
Veröffentlicht: (2024)
von: Li, Mingzhe, et al.
Veröffentlicht: (2024)
ClusterFusion++: Expanding Cluster-Level Fusion to Full Transformer-Block Decoding
von: Jin, ChiHeng, et al.
Veröffentlicht: (2026)
von: Jin, ChiHeng, et al.
Veröffentlicht: (2026)
SpComm3D: A Framework for Enabling Sparse Communication in 3D Sparse Kernels
von: Abubaker, Nabil, et al.
Veröffentlicht: (2024)
von: Abubaker, Nabil, et al.
Veröffentlicht: (2024)
Computational Grids
von: Foster, Ian, et al.
Veröffentlicht: (2025)
von: Foster, Ian, et al.
Veröffentlicht: (2025)
PACE Solver Description: twin_width_fmi
von: Balaban, David, et al.
Veröffentlicht: (2025)
von: Balaban, David, et al.
Veröffentlicht: (2025)
Empowering the Quantum Cloud User with QRIO
von: Chakraborty, Shmeelok, et al.
Veröffentlicht: (2024)
von: Chakraborty, Shmeelok, et al.
Veröffentlicht: (2024)
Exploring Dynamic Load Balancing Algorithms for Block-Structured Mesh-and-Particle Simulations in AMReX
von: Nanda, Amitash, et al.
Veröffentlicht: (2025)
von: Nanda, Amitash, et al.
Veröffentlicht: (2025)
MoLink: Distributed and Efficient Serving Framework for Large Models
von: Jin, Lewei, et al.
Veröffentlicht: (2025)
von: Jin, Lewei, et al.
Veröffentlicht: (2025)
SP-Chain: Boosting Intra-Shard and Cross-Shard Security and Performance in Blockchain Sharding
von: Li, Mingzhe, et al.
Veröffentlicht: (2024)
von: Li, Mingzhe, et al.
Veröffentlicht: (2024)
Distributed Ranges: A Model for Distributed Data Structures, Algorithms, and Views
von: Brock, Benjamin, et al.
Veröffentlicht: (2024)
von: Brock, Benjamin, et al.
Veröffentlicht: (2024)
SAGIPS: A Scalable Asynchronous Generative Inverse Problem Solver
von: Lersch, Daniel, et al.
Veröffentlicht: (2024)
von: Lersch, Daniel, et al.
Veröffentlicht: (2024)
Cache Blocking of Distributed-Memory Parallel Matrix Power Kernels
von: Lacey, Dane C., et al.
Veröffentlicht: (2024)
von: Lacey, Dane C., et al.
Veröffentlicht: (2024)
Energy Efficient Federated Learning with Hyperdimensional Computing (HDC)
von: Ding, Yahao, et al.
Veröffentlicht: (2026)
von: Ding, Yahao, et al.
Veröffentlicht: (2026)
Staging Blocked Evaluation over Structured Sparse Matrices
von: Das, Pratyush, et al.
Veröffentlicht: (2024)
von: Das, Pratyush, et al.
Veröffentlicht: (2024)
A Distributed-memory Tridiagonal Solver Based on a Specialised Data Structure Optimised for CPU and GPU Architectures
von: Akkurt, Semih, et al.
Veröffentlicht: (2024)
von: Akkurt, Semih, et al.
Veröffentlicht: (2024)
Highly Dynamic and Fully Distributed Data Structures
von: Augustine, John, et al.
Veröffentlicht: (2024)
von: Augustine, John, et al.
Veröffentlicht: (2024)
Min-Max Gathering on Infinite Grid
von: Chakraborty, Abhinav, et al.
Veröffentlicht: (2024)
von: Chakraborty, Abhinav, et al.
Veröffentlicht: (2024)
Multi-GPU Acceleration of PALABOS Fluid Solver using C++ Standard Parallelism
von: Latt, Jonas, et al.
Veröffentlicht: (2025)
von: Latt, Jonas, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CB-SpMV:A Data Aggregating and Balance Algorithm for Cache-Friendly Block-Based SpMV on GPUs
von: Cong, Xing, et al.
Veröffentlicht: (2026) -
Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage
von: Lin, Junqing, et al.
Veröffentlicht: (2025) -
MERBIT: A GPU-Based SpMV Method for Iterative Workloads
von: Zhang, Qi, et al.
Veröffentlicht: (2026) -
A Nonlinear Hash-based Optimization Method for SpMV on GPUs
von: Yan, Chen, et al.
Veröffentlicht: (2025) -
FlashMP: Fast Discrete Transform-Based Solver for Preconditioning Maxwell's Equations on GPUs
von: Zhang, Haoyuan, et al.
Veröffentlicht: (2025)