SpANNS: Optimizing Approximate Nearest Neighbor Search for Sparse Vectors Using Near Memory Processing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Tianqi, Ponzina, Flavio, Rosing, Tajana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FaTRQ: Tiered Residual Quantization for LLM Vector Search in Far-Memory-Aware ANNS Systems
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
Proxima: Near-storage Acceleration for Graph-based Approximate Nearest Neighbor Search in 3D NAND
von: Xu, Weihong, et al.
Veröffentlicht: (2023)
von: Xu, Weihong, et al.
Veröffentlicht: (2023)
HAVEN: High-Bandwidth Flash Augmented Vector Engine for Large-Scale Approximate Nearest-Neighbor Search Acceleration
von: Hsu, Po-Kai, et al.
Veröffentlicht: (2026)
von: Hsu, Po-Kai, et al.
Veröffentlicht: (2026)
HDDB: Efficient In-Storage SQL Database Search Using Hyperdimensional Computing on Ferroelectric NAND Flash
von: Zhao, Quanling, et al.
Veröffentlicht: (2025)
von: Zhao, Quanling, et al.
Veröffentlicht: (2025)
NasZip: Software and Hardware Co-Design to Accelerate Approximate Nearest Neighbor Search with DIMM-Based Near-Data Processing
von: Zou, Cheng, et al.
Veröffentlicht: (2026)
von: Zou, Cheng, et al.
Veröffentlicht: (2026)
Co-Designing Graph-based Approximate Nearest Neighbor Search at Billion Scale for Processing-in-Memory
von: Chen, Sitian, et al.
Veröffentlicht: (2026)
von: Chen, Sitian, et al.
Veröffentlicht: (2026)
Fast-OverlaPIM: A Fast Overlap-driven Mapping Framework for Processing In-Memory Neural Network Acceleration
von: Wang, Xuan, et al.
Veröffentlicht: (2024)
von: Wang, Xuan, et al.
Veröffentlicht: (2024)
NDSEARCH: Accelerating Graph-Traversal-Based Approximate Nearest Neighbor Search through Near Data Processing
von: Wang, Yitu, et al.
Veröffentlicht: (2023)
von: Wang, Yitu, et al.
Veröffentlicht: (2023)
Cosmos: A CXL-Based Full In-Memory System for Approximate Nearest Neighbor Search
von: Ko, Seoyoung, et al.
Veröffentlicht: (2025)
von: Ko, Seoyoung, et al.
Veröffentlicht: (2025)
FedUHD: Unsupervised Federated Learning using Hyperdimensional Computing
von: Lee, You Hak, et al.
Veröffentlicht: (2025)
von: Lee, You Hak, et al.
Veröffentlicht: (2025)
Efficient Open Modification Spectral Library Searching in High-Dimensional Space with Multi-Level-Cell Memory
von: Fan, Keming, et al.
Veröffentlicht: (2024)
von: Fan, Keming, et al.
Veröffentlicht: (2024)
pHNSW: PCA-Based Filtering to Accelerate HNSW Approximate Nearest Neighbor Search
von: Li, Zheng, et al.
Veröffentlicht: (2026)
von: Li, Zheng, et al.
Veröffentlicht: (2026)
SpecPCM: A Low-power PCM-based In-Memory Computing Accelerator for Full-stack Mass Spectrometry Analysis
von: Fan, Keming, et al.
Veröffentlicht: (2024)
von: Fan, Keming, et al.
Veröffentlicht: (2024)
Hybrid SLC-MLC RRAM Mixed-Signal Processing-in-Memory Architecture for Transformer Acceleration via Gradient Redistribution
von: Song, Chang Eun, et al.
Veröffentlicht: (2025)
von: Song, Chang Eun, et al.
Veröffentlicht: (2025)
CHIME: Chiplet-based Heterogeneous Near-Memory Acceleration for Edge Multimodal LLM Inference
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
GenDRAM:Hardware-Software Co-Design of General Platform in DRAM
von: Lu, Tsung-Han, et al.
Veröffentlicht: (2026)
von: Lu, Tsung-Han, et al.
Veröffentlicht: (2026)
GraphMatch: Subgraph Query Processing on FPGAs
von: Dann, Jonas, et al.
Veröffentlicht: (2024)
von: Dann, Jonas, et al.
Veröffentlicht: (2024)
PIM-FW: Hardware-Software Co-Design of All-pairs Shortest Paths in DRAM
von: Lu, Tsung-Han, et al.
Veröffentlicht: (2025)
von: Lu, Tsung-Han, et al.
Veröffentlicht: (2025)
UpANNS: Enhancing Billion-Scale ANNS Efficiency with Real-World PIM Architecture
von: Chen, Sitian, et al.
Veröffentlicht: (2024)
von: Chen, Sitian, et al.
Veröffentlicht: (2024)
Efficient Data Access Paths for Mixed Vector-Relational Search
von: Sanca, Viktor, et al.
Veröffentlicht: (2024)
von: Sanca, Viktor, et al.
Veröffentlicht: (2024)
RAPID-Graph: Recursive All-Pairs Shortest Paths Using Processing-in-Memory for Dynamic Programming on Graphs
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
PIMDAL: Mitigating the Memory Bottleneck in Data Analytics using a Real Processing-in-Memory System
von: Frouzakis, Manos, et al.
Veröffentlicht: (2025)
von: Frouzakis, Manos, et al.
Veröffentlicht: (2025)
Diba: A Re-configurable Stream Processor
von: Najafi, Mohammadreza, et al.
Veröffentlicht: (2023)
von: Najafi, Mohammadreza, et al.
Veröffentlicht: (2023)
Enthuse: Efficient Adaptable High-throughput Streaming Aggregation Engines
von: Papaphilippou, Philippos, et al.
Veröffentlicht: (2024)
von: Papaphilippou, Philippos, et al.
Veröffentlicht: (2024)
SwiftSpatial: Spatial Joins on Modern Hardware
von: Jiang, Wenqi, et al.
Veröffentlicht: (2023)
von: Jiang, Wenqi, et al.
Veröffentlicht: (2023)
Analyzing Performance Characteristics of PostgreSQL and MariaDB on NVMeVirt
von: Han, Juhee, et al.
Veröffentlicht: (2024)
von: Han, Juhee, et al.
Veröffentlicht: (2024)
FeNOMS: Enhancing Open Modification Spectral Library Search with In-Storage Processing on Ferroelectric NAND (FeNAND) Flash
von: Pinge, Sumukh, et al.
Veröffentlicht: (2025)
von: Pinge, Sumukh, et al.
Veröffentlicht: (2025)
SLIM: A Heterogeneous Accelerator for Edge Inference of Sparse Large Language Model via Adaptive Thresholding
von: Xu, Weihong, et al.
Veröffentlicht: (2025)
von: Xu, Weihong, et al.
Veröffentlicht: (2025)
Accelerating Multi-Scale Deformable Attention Using Near-Memory-Processing Architecture
von: Li, Huize, et al.
Veröffentlicht: (2026)
von: Li, Huize, et al.
Veröffentlicht: (2026)
Memory Hierarchy Design for Caching Middleware in the Age of NVM
von: Ghandeharizadeh, Shahram, et al.
Veröffentlicht: (2025)
von: Ghandeharizadeh, Shahram, et al.
Veröffentlicht: (2025)
JSPIM: A Skew-Aware PIM Accelerator for High-Performance Databases Join and Select Operations
von: Tajdari, Sabiha, et al.
Veröffentlicht: (2025)
von: Tajdari, Sabiha, et al.
Veröffentlicht: (2025)
RapidOMS: FPGA-based Open Modification Spectral Library Searching with HD Computing
von: Pinge, Sumukh, et al.
Veröffentlicht: (2024)
von: Pinge, Sumukh, et al.
Veröffentlicht: (2024)
Optimizing Structured-Sparse Matrix Multiplication in RISC-V Vector Processors
von: Titopoulos, Vasileios, et al.
Veröffentlicht: (2025)
von: Titopoulos, Vasileios, et al.
Veröffentlicht: (2025)
APACHE: A Processing-Near-Memory Architecture for Multi-Scheme Fully Homomorphic Encryption
von: Ding, Lin, et al.
Veröffentlicht: (2024)
von: Ding, Lin, et al.
Veröffentlicht: (2024)
Efficient Sparse Processing-in-Memory Architecture (ESPIM) for Machine Learning Inference
von: He, Mingxuan, et al.
Veröffentlicht: (2024)
von: He, Mingxuan, et al.
Veröffentlicht: (2024)
NVR: Vector Runahead on NPUs for Sparse Memory Access
von: Wang, Hui, et al.
Veröffentlicht: (2025)
von: Wang, Hui, et al.
Veröffentlicht: (2025)
Low-overhead General-purpose Near-Data Processing in CXL Memory Expanders
von: Ham, Hyungkyu, et al.
Veröffentlicht: (2024)
von: Ham, Hyungkyu, et al.
Veröffentlicht: (2024)
GEN-Graph: Heterogeneous PIM Accelerator for General Computational Patterns in Graph-based Dynamic Programming
von: Chen, Yanru, et al.
Veröffentlicht: (2026)
von: Chen, Yanru, et al.
Veröffentlicht: (2026)
Ironman: Accelerating Oblivious Transfer Extension for Privacy-Preserving AI with Near-Memory Processing
von: Lin, Chenqi, et al.
Veröffentlicht: (2025)
von: Lin, Chenqi, et al.
Veröffentlicht: (2025)
Qute: Towards Quantum-Native Database
von: Chen, Muzhi, et al.
Veröffentlicht: (2026)
von: Chen, Muzhi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
FaTRQ: Tiered Residual Quantization for LLM Vector Search in Far-Memory-Aware ANNS Systems
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026) -
Proxima: Near-storage Acceleration for Graph-based Approximate Nearest Neighbor Search in 3D NAND
von: Xu, Weihong, et al.
Veröffentlicht: (2023) -
HAVEN: High-Bandwidth Flash Augmented Vector Engine for Large-Scale Approximate Nearest-Neighbor Search Acceleration
von: Hsu, Po-Kai, et al.
Veröffentlicht: (2026) -
HDDB: Efficient In-Storage SQL Database Search Using Hyperdimensional Computing on Ferroelectric NAND Flash
von: Zhao, Quanling, et al.
Veröffentlicht: (2025) -
NasZip: Software and Hardware Co-Design to Accelerate Approximate Nearest Neighbor Search with DIMM-Based Near-Data Processing
von: Zou, Cheng, et al.
Veröffentlicht: (2026)