Breaking the Storage-Compute Bottleneck in Billion-Scale ANNS: A GPU-Driven Asynchronous I/O Framework
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiao, Yang, Sun, Mo, Song, Ziyu, Tian, Bing, Zhang, Jie, Sun, Jie, Wang, Zeke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FusionANNS: An Efficient CPU/GPU Cooperative Processing Architecture for Billion-scale Approximate Nearest Neighbor Search
von: Tian, Bing, et al.
Veröffentlicht: (2024)
von: Tian, Bing, et al.
Veröffentlicht: (2024)
Leveraging I/O Stalls for Efficient Scheduling in ANNS
von: Zhang, Juncheng, et al.
Veröffentlicht: (2026)
von: Zhang, Juncheng, et al.
Veröffentlicht: (2026)
GPU-Accelerated ANNS: Quantized for Speed, Built for Change
von: McCoy, Hunter, et al.
Veröffentlicht: (2026)
von: McCoy, Hunter, et al.
Veröffentlicht: (2026)
GRAB-ANNS: High-Throughput Indexing and Hybrid Search via GPU-Native Bucketing
von: Zhao, Xinkui, et al.
Veröffentlicht: (2026)
von: Zhao, Xinkui, et al.
Veröffentlicht: (2026)
CS-PQ: Cache-Friendly SIMD Product Quantization for Large-Scale ANNS Index Construction
von: Ma, Y. T., et al.
Veröffentlicht: (2026)
von: Ma, Y. T., et al.
Veröffentlicht: (2026)
Accelerating Graph Indexing for ANNS on Modern CPUs
von: Wang, Mengzhao, et al.
Veröffentlicht: (2025)
von: Wang, Mengzhao, et al.
Veröffentlicht: (2025)
StorageXTuner: An LLM Agent-Driven Automatic Tuning Framework for Heterogeneous Storage Systems
von: Lin, Qi, et al.
Veröffentlicht: (2025)
von: Lin, Qi, et al.
Veröffentlicht: (2025)
Bi-Directional Multi-Scale Graph Dataset Condensation via Information Bottleneck
von: Fu, Xingcheng, et al.
Veröffentlicht: (2024)
von: Fu, Xingcheng, et al.
Veröffentlicht: (2024)
SpANNS: Optimizing Approximate Nearest Neighbor Search for Sparse Vectors Using Near Memory Processing
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqi, et al.
Veröffentlicht: (2026)
Characterizing the Dilemma of Performance and Index Size in Billion-Scale Vector Search and Breaking It with Second-Tier Memory
von: Cheng, Rongxin, et al.
Veröffentlicht: (2024)
von: Cheng, Rongxin, et al.
Veröffentlicht: (2024)
SVFusion: A CPU-GPU Co-Processing Architecture for Large-Scale Real-Time Vector Search
von: Peng, Yuchen, et al.
Veröffentlicht: (2026)
von: Peng, Yuchen, et al.
Veröffentlicht: (2026)
UpANNS: Enhancing Billion-Scale ANNS Efficiency with Real-World PIM Architecture
von: Chen, Sitian, et al.
Veröffentlicht: (2024)
von: Chen, Sitian, et al.
Veröffentlicht: (2024)
Column-Oriented Datalog on the GPU
von: Sun, Yihao, et al.
Veröffentlicht: (2025)
von: Sun, Yihao, et al.
Veröffentlicht: (2025)
Effective and Efficient Conductance-based Community Search at Billion Scale
von: Lin, Longlong, et al.
Veröffentlicht: (2025)
von: Lin, Longlong, et al.
Veröffentlicht: (2025)
U-HNSW: An Efficient Graph-based Solution to ANNS Under Universal Lp Metrics
von: Wang, Huayi, et al.
Veröffentlicht: (2026)
von: Wang, Huayi, et al.
Veröffentlicht: (2026)
Asynchronous I/O -- With Great Power Comes Great Responsibility
von: Pestka, Constantin, et al.
Veröffentlicht: (2024)
von: Pestka, Constantin, et al.
Veröffentlicht: (2024)
OrchANN: A Unified I/O Orchestration Framework for Skewed Out-of-Core Vector Search
von: Huan, Chengying, et al.
Veröffentlicht: (2025)
von: Huan, Chengying, et al.
Veröffentlicht: (2025)
Optimizing Datalog for the GPU
von: Sun, Yihao, et al.
Veröffentlicht: (2023)
von: Sun, Yihao, et al.
Veröffentlicht: (2023)
SwitchDelta: Asynchronous Metadata Updating for Distributed Storage with In-Network Data Visibility
von: Li, Junru, et al.
Veröffentlicht: (2025)
von: Li, Junru, et al.
Veröffentlicht: (2025)
OSM+: Billion-Level OpenStreetMap Dataset for City-wide Experiments
von: Zheng, Guanjie, et al.
Veröffentlicht: (2025)
von: Zheng, Guanjie, et al.
Veröffentlicht: (2025)
I/O Optimizations for Graph-Based Disk-Resident Approximate Nearest Neighbor Search: A Design Space Exploration
von: Li, Liang, et al.
Veröffentlicht: (2026)
von: Li, Liang, et al.
Veröffentlicht: (2026)
Approximate Nearest Neighbor Search of Large Scale Vectors on Distributed Storage
von: Yu, Kun, et al.
Veröffentlicht: (2025)
von: Yu, Kun, et al.
Veröffentlicht: (2025)
PystachIO: Efficient Distributed GPU Query Processing with PyTorch over Fast Networks & Fast Storage
von: Luo, Jigao, et al.
Veröffentlicht: (2025)
von: Luo, Jigao, et al.
Veröffentlicht: (2025)
Revisiting the Design of In-Memory Dynamic Graph Storage
von: Su, Jixian, et al.
Veröffentlicht: (2025)
von: Su, Jixian, et al.
Veröffentlicht: (2025)
To GPU or Not to GPU: Vector Search in Relational Engines
von: Mageirakos, Vasilis, et al.
Veröffentlicht: (2026)
von: Mageirakos, Vasilis, et al.
Veröffentlicht: (2026)
B+ANN: A Fast Billion-Scale Disk-based Nearest-Neighbor Index
von: Tekin, Selim Furkan, et al.
Veröffentlicht: (2025)
von: Tekin, Selim Furkan, et al.
Veröffentlicht: (2025)
LLMATCH: A Unified Schema Matching Framework with Large Language Models
von: Wang, Sha, et al.
Veröffentlicht: (2025)
von: Wang, Sha, et al.
Veröffentlicht: (2025)
RapidStore: An Efficient Dynamic Graph Storage System for Concurrent Queries
von: Hao, Chiyu, et al.
Veröffentlicht: (2025)
von: Hao, Chiyu, et al.
Veröffentlicht: (2025)
A Cross-Chain Event-Driven Data Infrastructure for Aave Protocol Analytics and Applications
von: Fan, Junyi, et al.
Veröffentlicht: (2025)
von: Fan, Junyi, et al.
Veröffentlicht: (2025)
Rethinking Analytical Processing in the GPU Era
von: Yogatama, Bobbi, et al.
Veröffentlicht: (2025)
von: Yogatama, Bobbi, et al.
Veröffentlicht: (2025)
LSM-VEC: A Large-Scale Disk-Based System for Dynamic Vector Search
von: Zhong, Shurui, et al.
Veröffentlicht: (2025)
von: Zhong, Shurui, et al.
Veröffentlicht: (2025)
A GPU-Accelerated Framework for Multi-Attribute Range Filtered Approximate Nearest Neighbor Search
von: Li, Zhonggen, et al.
Veröffentlicht: (2026)
von: Li, Zhonggen, et al.
Veröffentlicht: (2026)
HDDB: Efficient In-Storage SQL Database Search Using Hyperdimensional Computing on Ferroelectric NAND Flash
von: Zhao, Quanling, et al.
Veröffentlicht: (2025)
von: Zhao, Quanling, et al.
Veröffentlicht: (2025)
3DPipe: A Pipelined GPU Framework for Scalable Generalized Spatial Join over Polyhedral Objects
von: Yuan, Lyuheng, et al.
Veröffentlicht: (2026)
von: Yuan, Lyuheng, et al.
Veröffentlicht: (2026)
Graph Structure Learning for Spatial-Temporal Imputation: Adapting to Node and Feature Scales
von: Yang, Xinyu, et al.
Veröffentlicht: (2024)
von: Yang, Xinyu, et al.
Veröffentlicht: (2024)
ACGraph: An Efficient Asynchronous Out-of-Core Graph Processing Framework
von: Chen, Dechuang, et al.
Veröffentlicht: (2025)
von: Chen, Dechuang, et al.
Veröffentlicht: (2025)
Quantifying Point Contributions: A Lightweight Framework for Efficient and Effective Query-Driven Trajectory Simplification
von: Song, Yumeng, et al.
Veröffentlicht: (2025)
von: Song, Yumeng, et al.
Veröffentlicht: (2025)
cuRPQ: A High-Performance GPU-Based Framework for Processing Regular and Conjunctive Regular Path Queries
von: Park, Sungwoo, et al.
Veröffentlicht: (2026)
von: Park, Sungwoo, et al.
Veröffentlicht: (2026)
LHGstore: An In-Memory Learned Graph Storage for Fast Updates and Analytics
von: Qiao, Pengpeng, et al.
Veröffentlicht: (2026)
von: Qiao, Pengpeng, et al.
Veröffentlicht: (2026)
A High-Throughput GPU Framework for Adaptive Lossless Compression of Floating-Point Data
von: Li, Zheng, et al.
Veröffentlicht: (2025)
von: Li, Zheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FusionANNS: An Efficient CPU/GPU Cooperative Processing Architecture for Billion-scale Approximate Nearest Neighbor Search
von: Tian, Bing, et al.
Veröffentlicht: (2024) -
Leveraging I/O Stalls for Efficient Scheduling in ANNS
von: Zhang, Juncheng, et al.
Veröffentlicht: (2026) -
GPU-Accelerated ANNS: Quantized for Speed, Built for Change
von: McCoy, Hunter, et al.
Veröffentlicht: (2026) -
GRAB-ANNS: High-Throughput Indexing and Hybrid Search via GPU-Native Bucketing
von: Zhao, Xinkui, et al.
Veröffentlicht: (2026) -
CS-PQ: Cache-Friendly SIMD Product Quantization for Large-Scale ANNS Index Construction
von: Ma, Y. T., et al.
Veröffentlicht: (2026)