PackSELL: A Sparse Matrix Format for Precision-Agnostic High-Performance SpMV
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Suzuki, Kengo, Iwashita, Takeshi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Systematic Literature Survey of Sparse Matrix-Vector Multiplication
von: Gao, Jianhua, et al.
Veröffentlicht: (2024)
von: Gao, Jianhua, et al.
Veröffentlicht: (2024)
Efficient Parallel Scheduling for Sparse Triangular Solvers
von: Böhnlein, Toni, et al.
Veröffentlicht: (2025)
von: Böhnlein, Toni, et al.
Veröffentlicht: (2025)
Data Scheduling Algorithm for Scalable and Efficient IoT Sensing in Cloud Computing
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
Precision-Aware Iterative Algorithms Based on Group-Shared Exponents of Floating-Point Numbers
von: Gao, Jianhua, et al.
Veröffentlicht: (2024)
von: Gao, Jianhua, et al.
Veröffentlicht: (2024)
Cascaded Prediction and Asynchronous Execution of Iterative Algorithms on Heterogeneous Platforms
von: Gao, Jianhua, et al.
Veröffentlicht: (2024)
von: Gao, Jianhua, et al.
Veröffentlicht: (2024)
Arrow Matrix Decomposition: A Novel Approach for Communication-Efficient Sparse Matrix Multiplication
von: Gianinazzi, Lukas, et al.
Veröffentlicht: (2024)
von: Gianinazzi, Lukas, et al.
Veröffentlicht: (2024)
Synthesis of signal processing algorithms with constraints on minimal parallelism and memory space
von: Salishev, Sergey
Veröffentlicht: (2025)
von: Salishev, Sergey
Veröffentlicht: (2025)
The Impact of Partial Computations on the Red-Blue Pebble Game
von: Papp, Pál András, et al.
Veröffentlicht: (2025)
von: Papp, Pál András, et al.
Veröffentlicht: (2025)
Low-Bandwidth Matrix Multiplication: Faster Algorithms and More General Forms of Sparsity
von: Gupta, Chetan, et al.
Veröffentlicht: (2024)
von: Gupta, Chetan, et al.
Veröffentlicht: (2024)
Replication in Graph Partitioning and Scheduling Problems
von: Papp, Pál András, et al.
Veröffentlicht: (2026)
von: Papp, Pál András, et al.
Veröffentlicht: (2026)
Understanding GEMM Performance and Energy on NVIDIA Ada Lovelace: A Machine Learning-Based Analytical Approach
von: Xiaoteng, et al.
Veröffentlicht: (2024)
von: Xiaoteng, et al.
Veröffentlicht: (2024)
Conflict-Freedom as a Progress Condition
von: Kuznetsov, Petr, et al.
Veröffentlicht: (2026)
von: Kuznetsov, Petr, et al.
Veröffentlicht: (2026)
Red-Blue Pebbling with Multiple Processors: Time, Communication and Memory Trade-offs
von: Böhnlein, Toni, et al.
Veröffentlicht: (2024)
von: Böhnlein, Toni, et al.
Veröffentlicht: (2024)
Declarative distributed algorithms as axiomatic theories in three-valued modal logic over semitopologies
von: Gabbay, Murdoch J.
Veröffentlicht: (2025)
von: Gabbay, Murdoch J.
Veröffentlicht: (2025)
Acc-SpMM: Accelerating General-purpose Sparse Matrix-Matrix Multiplication with GPU Tensor Cores
von: Zhao, Haisha, et al.
Veröffentlicht: (2025)
von: Zhao, Haisha, et al.
Veröffentlicht: (2025)
Gradient Coding with Iterative Block Leverage Score Sampling
von: Charalambides, Neophytos, et al.
Veröffentlicht: (2023)
von: Charalambides, Neophytos, et al.
Veröffentlicht: (2023)
Accelerating State-Vector Quantum Simulation on Integrated GPUs via Cache Locality Optimization: A Cross-Architecture Evaluation
von: Thomaz, Gabriel Fernandes, et al.
Veröffentlicht: (2026)
von: Thomaz, Gabriel Fernandes, et al.
Veröffentlicht: (2026)
Distributed Hybrid Sketching for $\ell_2$-Embeddings
von: Charalambides, Neophytos, et al.
Veröffentlicht: (2024)
von: Charalambides, Neophytos, et al.
Veröffentlicht: (2024)
A simple protocol to automate the executing, scaling, and reconfiguration of Cloud-Native Apps
von: Ambroszkiewicz, Stanislaw, et al.
Veröffentlicht: (2023)
von: Ambroszkiewicz, Stanislaw, et al.
Veröffentlicht: (2023)
Supercomputers as a Continous Medium
von: Karp, Martin, et al.
Veröffentlicht: (2024)
von: Karp, Martin, et al.
Veröffentlicht: (2024)
Solving Large Rank-Deficient Linear Least-Squares Problems on Shared-Memory CPU Architectures and GPU Architectures
von: Chillarón, Mónica, et al.
Veröffentlicht: (2024)
von: Chillarón, Mónica, et al.
Veröffentlicht: (2024)
Accelerating Matrix Multiplication: A Performance Comparison Between Multi-Core CPU and GPU
von: Ansari, Mufakir Qamar, et al.
Veröffentlicht: (2025)
von: Ansari, Mufakir Qamar, et al.
Veröffentlicht: (2025)
Parallel Quadratic Selected Inversion in Quantum Transport Simulation
von: Maillou, Vincent, et al.
Veröffentlicht: (2026)
von: Maillou, Vincent, et al.
Veröffentlicht: (2026)
Pure Data Spaces
von: Youssef, Saul
Veröffentlicht: (2025)
von: Youssef, Saul
Veröffentlicht: (2025)
Decidability Issues for Petri Nets -- a survey
von: Esparza, Javier, et al.
Veröffentlicht: (2024)
von: Esparza, Javier, et al.
Veröffentlicht: (2024)
Stochastic well-structured transition systems
von: Aspnes, James
Veröffentlicht: (2025)
von: Aspnes, James
Veröffentlicht: (2025)
Parallelization Strategies for the Randomized Kaczmarz Algorithm on Large-Scale Dense Systems
von: Ferreira, Inês, et al.
Veröffentlicht: (2024)
von: Ferreira, Inês, et al.
Veröffentlicht: (2024)
Approximate Distributed Coded Computing: Polynomial Codes and Randomized Sketching
von: Charalambides, Neophytos, et al.
Veröffentlicht: (2026)
von: Charalambides, Neophytos, et al.
Veröffentlicht: (2026)
Serial Parallel Reliability Redundancy Allocation Optimization for Energy Efficient and Fault Tolerant Cloud Computing
von: Krishna, Gutha Jaya
Veröffentlicht: (2024)
von: Krishna, Gutha Jaya
Veröffentlicht: (2024)
Design, Configuration, Implementation, and Performance of a Simple 32 Core Raspberry Pi Cluster
von: Cicirello, Vincent A.
Veröffentlicht: (2017)
von: Cicirello, Vincent A.
Veröffentlicht: (2017)
Computational Power of Opaque Robots
von: Feletti, Caterina, et al.
Veröffentlicht: (2024)
von: Feletti, Caterina, et al.
Veröffentlicht: (2024)
An Empirical Evaluation of Quantum-Inspired QUBO Methods for Heterogeneous HPC Workflow Mapping and Scheduling
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2026)
von: Sharma, Aasish Kumar, et al.
Veröffentlicht: (2026)
Exploiting nested task-parallelism in the $\mathcal{H}-LU$ factorization
von: Carratalá-Sáez, Rocío, et al.
Veröffentlicht: (2019)
von: Carratalá-Sáez, Rocío, et al.
Veröffentlicht: (2019)
Efficiently Scheduling Parallel DAG Tasks on Identical Multiprocessors
von: Lendve, Shardul, et al.
Veröffentlicht: (2024)
von: Lendve, Shardul, et al.
Veröffentlicht: (2024)
Dynamic Memory Management on GPUs with SYCL
von: Standish, Russell K.
Veröffentlicht: (2025)
von: Standish, Russell K.
Veröffentlicht: (2025)
Floating Point Compression of Hierarchical Matrix Formats and its Impact on Matrix-Vector Multiplication
von: Kriemann, Ronald
Veröffentlicht: (2024)
von: Kriemann, Ronald
Veröffentlicht: (2024)
Computing Inductive Invariants of Regular Abstraction Frameworks
von: Czerner, Philipp, et al.
Veröffentlicht: (2024)
von: Czerner, Philipp, et al.
Veröffentlicht: (2024)
Graded modal logic and counting message passing automata
von: Ahvonen, Veeti, et al.
Veröffentlicht: (2024)
von: Ahvonen, Veeti, et al.
Veröffentlicht: (2024)
On the Computation of 2-Dimensional Recurrence Equations
von: Natale, Giuseppe
Veröffentlicht: (2024)
von: Natale, Giuseppe
Veröffentlicht: (2024)
Scalable Dual Coordinate Descent for Kernel Methods
von: Shao, Zishan, et al.
Veröffentlicht: (2024)
von: Shao, Zishan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Systematic Literature Survey of Sparse Matrix-Vector Multiplication
von: Gao, Jianhua, et al.
Veröffentlicht: (2024) -
Efficient Parallel Scheduling for Sparse Triangular Solvers
von: Böhnlein, Toni, et al.
Veröffentlicht: (2025) -
Data Scheduling Algorithm for Scalable and Efficient IoT Sensing in Cloud Computing
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025) -
Precision-Aware Iterative Algorithms Based on Group-Shared Exponents of Floating-Point Numbers
von: Gao, Jianhua, et al.
Veröffentlicht: (2024) -
Cascaded Prediction and Asynchronous Execution of Iterative Algorithms on Heterogeneous Platforms
von: Gao, Jianhua, et al.
Veröffentlicht: (2024)