RSR-core: A High-Performance Engine for Low-Bit Matrix-Vector Multiplication
Fuente:
arXiv
Saved in:
| Main Authors: | Dehghankar, Mohsen, Asudeh, Abolfazl |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Efficient Matrix Multiplication Algorithm for Accelerating Inference in Binary and Ternary Neural Networks
by: Dehghankar, Mohsen, et al.
Published: (2024)
by: Dehghankar, Mohsen, et al.
Published: (2024)
HENN: A Hierarchical Epsilon Net Navigation Graph for Approximate Nearest Neighbor Search
by: Dehghankar, Mohsen, et al.
Published: (2025)
by: Dehghankar, Mohsen, et al.
Published: (2025)
On Fair Epsilon Net and Geometric Hitting Set
by: Dehghankar, Mohsen, et al.
Published: (2025)
by: Dehghankar, Mohsen, et al.
Published: (2025)
Fair Set Cover
by: Dehghankar, Mohsen, et al.
Published: (2024)
by: Dehghankar, Mohsen, et al.
Published: (2024)
Revisiting Forest Proximities via Sparse Leaf-Incidence Kernels
by: Aumon, Adrien, et al.
Published: (2026)
by: Aumon, Adrien, et al.
Published: (2026)
The Transient Cost of Learning in Queueing Systems
by: Freund, Daniel, et al.
Published: (2023)
by: Freund, Daniel, et al.
Published: (2023)
Toward Greener Matrix Operations by Lossless Compressed Formats
by: Tosoni, Francesco, et al.
Published: (2024)
by: Tosoni, Francesco, et al.
Published: (2024)
A Simple Sparse Matrix Vector Multiplication Approach to Padded Convolution
by: Chaudhry, Zan
Published: (2024)
by: Chaudhry, Zan
Published: (2024)
Improved Algorithms for Kernel Matrix-Vector Multiplication Under Sparsity Assumptions
by: Indyk, Piotr, et al.
Published: (2025)
by: Indyk, Piotr, et al.
Published: (2025)
Dynamic Necklace Splitting
by: Advani, Rishi, et al.
Published: (2025)
by: Advani, Rishi, et al.
Published: (2025)
Fair-Count-Min: Frequency Estimation under Equal Group-wise Approximation Factor
by: Shahbazi, Nima, et al.
Published: (2025)
by: Shahbazi, Nima, et al.
Published: (2025)
A Fair and Memory/Time-efficient Hashmap
by: Asudeh, Abolfazl, et al.
Published: (2023)
by: Asudeh, Abolfazl, et al.
Published: (2023)
Optimal Sketching for Residual Error Estimation for Matrix and Vector Norms
by: Li, Yi, et al.
Published: (2024)
by: Li, Yi, et al.
Published: (2024)
Cheesemap: A High-Performance Point-Indexing Data Structure for Neighbor Search in LiDAR Data
by: Laso, Ruben, et al.
Published: (2025)
by: Laso, Ruben, et al.
Published: (2025)
Diagonally-Addressed Matrix Nicknack: How to improve SpMV performance
by: Saak, Jens, et al.
Published: (2023)
by: Saak, Jens, et al.
Published: (2023)
Optimal Approximate Matrix Multiplication over Sliding Windows
by: Yao, Ziqi, et al.
Published: (2025)
by: Yao, Ziqi, et al.
Published: (2025)
Elastic Sketch under Random Stationary Streams: Limiting Behavior and Near-Optimal Configuration
by: Mazziane, Younes Ben, et al.
Published: (2026)
by: Mazziane, Younes Ben, et al.
Published: (2026)
Virtual-Memory Powersort
by: Moltmann, Finn, et al.
Published: (2026)
by: Moltmann, Finn, et al.
Published: (2026)
FRSZ2 for In-Register Block Compression Inside GMRES on GPUs
by: Grützmacher, Thomas, et al.
Published: (2024)
by: Grützmacher, Thomas, et al.
Published: (2024)
Accurate and Fast Approximate Graph Pattern Mining at Scale
by: Arpaci-Dusseau, Anna, et al.
Published: (2024)
by: Arpaci-Dusseau, Anna, et al.
Published: (2024)
Less is More: Faster Maximum Clique Search by Work-Avoidance
by: Vandierendonck, Hans
Published: (2025)
by: Vandierendonck, Hans
Published: (2025)
Count-Min Sketch with Conservative Updates: Worst-Case Analysis
by: Mazziane, Younes Ben, et al.
Published: (2024)
by: Mazziane, Younes Ben, et al.
Published: (2024)
HashEvict: A Pre-Attention KV Cache Eviction Strategy using Locality-Sensitive Hashing
by: Liu, Minghui, et al.
Published: (2024)
by: Liu, Minghui, et al.
Published: (2024)
MACKO: Sparse Matrix-Vector Multiplication for Low Sparsity
by: Macko, Vladimír, et al.
Published: (2025)
by: Macko, Vladimír, et al.
Published: (2025)
Toward Efficient and Scalable Design of In-Memory Graph-Based Vector Search
by: Azizi, Ilias, et al.
Published: (2025)
by: Azizi, Ilias, et al.
Published: (2025)
FB$^+$-tree: A Memory-Optimized B$^+$-tree with Latch-Free Update
by: Chen, Yuan, et al.
Published: (2025)
by: Chen, Yuan, et al.
Published: (2025)
CAMP: A Cost Adaptive Multi-Queue Eviction Policy for Key-Value Stores
by: Ghandeharizadeh, Shahram, et al.
Published: (2024)
by: Ghandeharizadeh, Shahram, et al.
Published: (2024)
Rank It, Then Ask It: Input Reranking for Maximizing the Performance of LLMs on Symmetric Tasks
by: Dehghankar, Mohsen, et al.
Published: (2024)
by: Dehghankar, Mohsen, et al.
Published: (2024)
PHast -- Perfect Hashing made fast
by: Beling, Piotr, et al.
Published: (2025)
by: Beling, Piotr, et al.
Published: (2025)
Adaptive Hybrid Sort: Dynamic Strategy Selection for Optimal Sorting Across Diverse Data Distributions
by: Balasubramanian, Shrinivass Arunachalam
Published: (2025)
by: Balasubramanian, Shrinivass Arunachalam
Published: (2025)
Online Maximum Independent Set of Hyperrectangles
by: Advani, Rishi, et al.
Published: (2023)
by: Advani, Rishi, et al.
Published: (2023)
Scheduling with Uncertain Holding Costs and its Application to Content Moderation
by: Gocmen, Caner, et al.
Published: (2025)
by: Gocmen, Caner, et al.
Published: (2025)
Results of the Big ANN: NeurIPS'23 competition
by: Simhadri, Harsha Vardhan, et al.
Published: (2024)
by: Simhadri, Harsha Vardhan, et al.
Published: (2024)
Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache
by: Dehghankar, Mohsen, et al.
Published: (2026)
by: Dehghankar, Mohsen, et al.
Published: (2026)
Mining the Minoria: Unknown, Under-represented, and Under-performing Minority Groups
by: Dehghankar, Mohsen, et al.
Published: (2024)
by: Dehghankar, Mohsen, et al.
Published: (2024)
bsort: A theoretically efficient non-comparison-based sorting algorithm for integer and floating-point numbers
by: Guzmán, Benjamín
Published: (2026)
by: Guzmán, Benjamín
Published: (2026)
Beyond Worst-Case Dimensionality Reduction for Sparse Vectors
by: Silwal, Sandeep, et al.
Published: (2025)
by: Silwal, Sandeep, et al.
Published: (2025)
OpenTensor: Reproducing Faster Matrix Multiplication Discovering Algorithms
by: Sun, Yiwen, et al.
Published: (2024)
by: Sun, Yiwen, et al.
Published: (2024)
Tight Differentially Private PCA via Matrix Coherence
by: d'Orsi, Tommaso, et al.
Published: (2025)
by: d'Orsi, Tommaso, et al.
Published: (2025)
GPU Acceleration of Sparse Fully Homomorphic Encrypted DNNs
by: D'Agata, Lara, et al.
Published: (2026)
by: D'Agata, Lara, et al.
Published: (2026)
Similar Items
-
An Efficient Matrix Multiplication Algorithm for Accelerating Inference in Binary and Ternary Neural Networks
by: Dehghankar, Mohsen, et al.
Published: (2024) -
HENN: A Hierarchical Epsilon Net Navigation Graph for Approximate Nearest Neighbor Search
by: Dehghankar, Mohsen, et al.
Published: (2025) -
On Fair Epsilon Net and Geometric Hitting Set
by: Dehghankar, Mohsen, et al.
Published: (2025) -
Fair Set Cover
by: Dehghankar, Mohsen, et al.
Published: (2024) -
Revisiting Forest Proximities via Sparse Leaf-Incidence Kernels
by: Aumon, Adrien, et al.
Published: (2026)