BANG: Billion-Scale Approximate Nearest Neighbor Search using a Single GPU
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | V., Karthik, Khan, Saim, Singh, Somesh, Simhadri, Harsha Vardhan, Vedurada, Jyothi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PathWeaver: A High-Throughput Multi-GPU System for Graph-Based Approximate Nearest Neighbor Search
von: Kim, Sukjin, et al.
Veröffentlicht: (2025)
von: Kim, Sukjin, et al.
Veröffentlicht: (2025)
Scalable Graph Indexing using GPUs for Approximate Nearest Neighbor Search
von: Li, Zhonggen, et al.
Veröffentlicht: (2025)
von: Li, Zhonggen, et al.
Veröffentlicht: (2025)
GRNND: A GPU-Parallel Relative NN-Descent Algorithm for Efficient Approximate Nearest Neighbor Graph Construction
von: Li, Xiang, et al.
Veröffentlicht: (2025)
von: Li, Xiang, et al.
Veröffentlicht: (2025)
Arkade: k-Nearest Neighbor Search With Non-Euclidean Distances using GPU Ray Tracing
von: Mandarapu, Durga, et al.
Veröffentlicht: (2023)
von: Mandarapu, Durga, et al.
Veröffentlicht: (2023)
Efficient Graph-Based Approximate Nearest Neighbor Search Achieving: Low Latency Without Throughput Loss
von: Luo, Jingjia, et al.
Veröffentlicht: (2025)
von: Luo, Jingjia, et al.
Veröffentlicht: (2025)
Advancing RT Core-Accelerated Fixed-Radius Nearest Neighbor Search
von: Meneses, Enzo, et al.
Veröffentlicht: (2026)
von: Meneses, Enzo, et al.
Veröffentlicht: (2026)
Exact Nearest-Neighbor Search on Energy-Efficient FPGA Devices
von: Dazzi, Patrizio, et al.
Veröffentlicht: (2025)
von: Dazzi, Patrizio, et al.
Veröffentlicht: (2025)
On the Effectiveness of Graph Reordering for Accelerating Approximate Nearest Neighbor Search on GPU
von: Oguri, Yutaro, et al.
Veröffentlicht: (2025)
von: Oguri, Yutaro, et al.
Veröffentlicht: (2025)
SOLANET: Distributed Neighbor Graph Construction on GPU-Accelerated Systems
von: Iwabuchi, Keita, et al.
Veröffentlicht: (2026)
von: Iwabuchi, Keita, et al.
Veröffentlicht: (2026)
Optimizing Distributed Training Approaches for Scaling Neural Networks
von: Baligodugula, Vishnu Vardhan, et al.
Veröffentlicht: (2025)
von: Baligodugula, Vishnu Vardhan, et al.
Veröffentlicht: (2025)
PECANN: Parallel Efficient Clustering with Graph-Based Approximate Nearest Neighbor Search
von: Yu, Shangdi, et al.
Veröffentlicht: (2023)
von: Yu, Shangdi, et al.
Veröffentlicht: (2023)
NasZip: Software and Hardware Co-Design to Accelerate Approximate Nearest Neighbor Search with DIMM-Based Near-Data Processing
von: Zou, Cheng, et al.
Veröffentlicht: (2026)
von: Zou, Cheng, et al.
Veröffentlicht: (2026)
ATLAS: Efficient Out-of-Core Inference for Billion-Scale Graph Neural Networks
von: Naman, Pranjal, et al.
Veröffentlicht: (2026)
von: Naman, Pranjal, et al.
Veröffentlicht: (2026)
AMPED: Accelerating MTTKRP for Billion-Scale Sparse Tensor Decomposition on Multiple GPUs
von: Wijeratne, Sasindu, et al.
Veröffentlicht: (2025)
von: Wijeratne, Sasindu, et al.
Veröffentlicht: (2025)
AQUA: Network-Accelerated Memory Offloading for LLMs in Scale-Up GPU Domains
von: Kumar, Abhishek Vijaya, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek Vijaya, et al.
Veröffentlicht: (2024)
CleANN: Efficient Full Dynamism in Graph-based Approximate Nearest Neighbor Search
von: Zhang, Ziyu, et al.
Veröffentlicht: (2025)
von: Zhang, Ziyu, et al.
Veröffentlicht: (2025)
PilotANN: Memory-Bounded GPU Acceleration for Vector Search
von: Gui, Yuntao, et al.
Veröffentlicht: (2025)
von: Gui, Yuntao, et al.
Veröffentlicht: (2025)
ParvaGPU: Efficient Spatial GPU Sharing for Large-Scale DNN Inference in Cloud Environments
von: Lee, Munkyu, et al.
Veröffentlicht: (2024)
von: Lee, Munkyu, et al.
Veröffentlicht: (2024)
GPU-Accelerated Vecchia Approximations of Gaussian Processes for Geospatial Data using Batched Matrix Computations
von: Pan, Qilong, et al.
Veröffentlicht: (2024)
von: Pan, Qilong, et al.
Veröffentlicht: (2024)
Fantasy: Efficient Large-scale Vector Search on GPU Clusters with GPUDirect Async
von: Liu, Yi, et al.
Veröffentlicht: (2025)
von: Liu, Yi, et al.
Veröffentlicht: (2025)
PiPNN: Ultra-Scalable Graph-Based Nearest Neighbor Indexing
von: Rubel, Tobias, et al.
Veröffentlicht: (2026)
von: Rubel, Tobias, et al.
Veröffentlicht: (2026)
FastTrack: GPU-Accelerated Tracking for Visual SLAM
von: Khabiri, Kimia, et al.
Veröffentlicht: (2025)
von: Khabiri, Kimia, et al.
Veröffentlicht: (2025)
Poplar: Efficient Scaling of Distributed DNN Training on Heterogeneous GPU Clusters
von: Zhang, WenZheng, et al.
Veröffentlicht: (2024)
von: Zhang, WenZheng, et al.
Veröffentlicht: (2024)
Efficient Graph Embedding at Scale: Optimizing CPU-GPU-SSD Integration
von: Li, Zhonggen, et al.
Veröffentlicht: (2025)
von: Li, Zhonggen, et al.
Veröffentlicht: (2025)
Heimdall++: Optimizing GPU Utilization and Pipeline Parallelism for Efficient Single-Pulse Detection
von: Xia, Bingzheng, et al.
Veröffentlicht: (2025)
von: Xia, Bingzheng, et al.
Veröffentlicht: (2025)
PRISM: Dynamic Primitive-Based Forecasting for Large-Scale GPU Cluster Workloads
von: Wu, Xin, et al.
Veröffentlicht: (2026)
von: Wu, Xin, et al.
Veröffentlicht: (2026)
City-Scale Visibility Graph Analysis via GPU-Accelerated HyperBall
von: Hodge, Alex, et al.
Veröffentlicht: (2026)
von: Hodge, Alex, et al.
Veröffentlicht: (2026)
KIS-S: A GPU-Aware Kubernetes Inference Simulator with RL-Based Auto-Scaling
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
von: Zhang, Guilin, et al.
Veröffentlicht: (2025)
Beyond Microservices: Testing Web-Scale RCA Methods on GPU-Driven LLM Workloads
von: Scheinert, Dominik, et al.
Veröffentlicht: (2026)
von: Scheinert, Dominik, et al.
Veröffentlicht: (2026)
Characterizing the Dilemma of Performance and Index Size in Billion-Scale Vector Search and Breaking It with Second-Tier Memory
von: Cheng, Rongxin, et al.
Veröffentlicht: (2024)
von: Cheng, Rongxin, et al.
Veröffentlicht: (2024)
DISTRIBUTEDANN: Efficient Scaling of a Single DISKANN Graph Across Thousands of Computers
von: Adams, Philip, et al.
Veröffentlicht: (2025)
von: Adams, Philip, et al.
Veröffentlicht: (2025)
Large Scale Multi-GPU Based Parallel Traffic Simulation for Accelerated Traffic Assignment and Propagation
von: Jiang, Xuan, et al.
Veröffentlicht: (2024)
von: Jiang, Xuan, et al.
Veröffentlicht: (2024)
Intel(R) SHMEM: GPU-initiated OpenSHMEM using SYCL
von: Brooks, Alex, et al.
Veröffentlicht: (2024)
von: Brooks, Alex, et al.
Veröffentlicht: (2024)
Characterizing Production GPU Workloads using System-wide Telemetry Data
von: Cankur, Onur, et al.
Veröffentlicht: (2025)
von: Cankur, Onur, et al.
Veröffentlicht: (2025)
CAGRA: Highly Parallel Graph Construction and Approximate Nearest Neighbor Search for GPUs
von: Ootomo, Hiroyuki, et al.
Veröffentlicht: (2023)
von: Ootomo, Hiroyuki, et al.
Veröffentlicht: (2023)
Understanding GPU Triggering APIs for MPI+X Communication
von: Bridges, Patrick G., et al.
Veröffentlicht: (2024)
von: Bridges, Patrick G., et al.
Veröffentlicht: (2024)
Improving GPU Multi-Tenancy Through Dynamic Multi-Instance GPU Reconfiguration
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
Multi-GPU Acceleration of PALABOS Fluid Solver using C++ Standard Parallelism
von: Latt, Jonas, et al.
Veröffentlicht: (2025)
von: Latt, Jonas, et al.
Veröffentlicht: (2025)
Fast Iterative Graph Computing with Updated Neighbor States
von: Zhou, Yijie, et al.
Veröffentlicht: (2024)
von: Zhou, Yijie, et al.
Veröffentlicht: (2024)
Scaling Up Throughput-oriented LLM Inference Applications on Heterogeneous Opportunistic GPU Clusters with Pervasive Context Management
von: Phung, Thanh Son, et al.
Veröffentlicht: (2025)
von: Phung, Thanh Son, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PathWeaver: A High-Throughput Multi-GPU System for Graph-Based Approximate Nearest Neighbor Search
von: Kim, Sukjin, et al.
Veröffentlicht: (2025) -
Scalable Graph Indexing using GPUs for Approximate Nearest Neighbor Search
von: Li, Zhonggen, et al.
Veröffentlicht: (2025) -
GRNND: A GPU-Parallel Relative NN-Descent Algorithm for Efficient Approximate Nearest Neighbor Graph Construction
von: Li, Xiang, et al.
Veröffentlicht: (2025) -
Arkade: k-Nearest Neighbor Search With Non-Euclidean Distances using GPU Ray Tracing
von: Mandarapu, Durga, et al.
Veröffentlicht: (2023) -
Efficient Graph-Based Approximate Nearest Neighbor Search Achieving: Low Latency Without Throughput Loss
von: Luo, Jingjia, et al.
Veröffentlicht: (2025)