Efficient and Reliable Vector Similarity Search Using Asymmetric Encoding with NAND-Flash for Many-Class Few-Shot Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Chiang, Hao-Wei, Huang, Chi-Tse, Cheng, Hsiang-Yun, Tseng, Po-Hao, Lee, Ming-Hsiu, An-Yeu, Wu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Search-in-Memory (SiM): Reliable, Versatile, and Efficient Data Matching in SSD's NAND Flash Memory Chip for Data Indexing Acceleration
por: Chen, Yun-Chih, et al.
Publicado: (2024)
por: Chen, Yun-Chih, et al.
Publicado: (2024)
FeNOMS: Enhancing Open Modification Spectral Library Search with In-Storage Processing on Ferroelectric NAND (FeNAND) Flash
por: Pinge, Sumukh, et al.
Publicado: (2025)
por: Pinge, Sumukh, et al.
Publicado: (2025)
Flexible In-NAND Cryptographic Processing for Secure Flash Storage
por: Noh, Seock-Hwan, et al.
Publicado: (2025)
por: Noh, Seock-Hwan, et al.
Publicado: (2025)
HDDB: Efficient In-Storage SQL Database Search Using Hyperdimensional Computing on Ferroelectric NAND Flash
por: Zhao, Quanling, et al.
Publicado: (2025)
por: Zhao, Quanling, et al.
Publicado: (2025)
AERO: Adaptive Erase Operation for Improving Lifetime and Performance of Modern NAND Flash-Based SSDs
por: Cho, Sungjun, et al.
Publicado: (2024)
por: Cho, Sungjun, et al.
Publicado: (2024)
Proxima: Near-storage Acceleration for Graph-based Approximate Nearest Neighbor Search in 3D NAND
por: Xu, Weihong, et al.
Publicado: (2023)
por: Xu, Weihong, et al.
Publicado: (2023)
Dissecting and Re-architecting 3D NAND Flash PIM Arrays for Efficient Single-Batch Token Generation in LLMs
por: Jang, Yongjoo, et al.
Publicado: (2025)
por: Jang, Yongjoo, et al.
Publicado: (2025)
STRAW: A Stress-Aware WL-Based Read Reclaim Technique for High-Density NAND Flash-Based SSDs
por: Chun, Myoungjun, et al.
Publicado: (2025)
por: Chun, Myoungjun, et al.
Publicado: (2025)
MCFlash: Bulk Bitwise Processing in 3D NAND with Dynamic Sensing and Multi-level Encoding
por: Rahman, Habib Ur, et al.
Publicado: (2026)
por: Rahman, Habib Ur, et al.
Publicado: (2026)
HAVEN: High-Bandwidth Flash Augmented Vector Engine for Large-Scale Approximate Nearest-Neighbor Search Acceleration
por: Hsu, Po-Kai, et al.
Publicado: (2026)
por: Hsu, Po-Kai, et al.
Publicado: (2026)
NVLLM: A 3D NAND-Centric Architecture Enabling Edge on-Device LLM Inference
por: Hao, Mingbo, et al.
Publicado: (2026)
por: Hao, Mingbo, et al.
Publicado: (2026)
NASiC: 3D NAND-based CAM-Selected Multibit CIM Architecture for Efficient On-Device Mixture-of-Experts LLM Inference
por: Xu, Weikai, et al.
Publicado: (2026)
por: Xu, Weikai, et al.
Publicado: (2026)
Design Environment of Quantization-Aware Edge AI Hardware for Few-Shot Learning
por: Kanda, R., et al.
Publicado: (2026)
por: Kanda, R., et al.
Publicado: (2026)
Hardware-Aware Neural Dropout Search for Reliable Uncertainty Prediction on FPGA
por: Zhang, Zehuan, et al.
Publicado: (2024)
por: Zhang, Zehuan, et al.
Publicado: (2024)
Bit-Width-Aware Design Environment for Few-Shot Learning on Edge AI Hardware
por: Kanda, R., et al.
Publicado: (2026)
por: Kanda, R., et al.
Publicado: (2026)
AutoPower: Automated Few-Shot Architecture-Level Power Modeling by Power Group Decoupling
por: Zhang, Qijun, et al.
Publicado: (2025)
por: Zhang, Qijun, et al.
Publicado: (2025)
Fast Graph Vector Search via Hardware Acceleration and Delayed-Synchronization Traversal
por: Jiang, Wenqi, et al.
Publicado: (2024)
por: Jiang, Wenqi, et al.
Publicado: (2024)
From Characterization to Microarchitecture: Designing an Elegant and Reliable BFP-Based NPU
por: Zhang, Jie, et al.
Publicado: (2026)
por: Zhang, Jie, et al.
Publicado: (2026)
A Tensor-Train Decomposition based Compression of LLMs on Group Vector Systolic Accelerator
por: Huang, Sixiao, et al.
Publicado: (2025)
por: Huang, Sixiao, et al.
Publicado: (2025)
Systolic Sparse Tensor Slices: FPGA Building Blocks for Sparse and Dense AI Acceleration
por: Taka, Endri, et al.
Publicado: (2025)
por: Taka, Endri, et al.
Publicado: (2025)
Towards Efficient and Accurate Detection of On-Chip Fail-Slow Failures for Many-Core Accelerators
por: Wu, Junchi, et al.
Publicado: (2025)
por: Wu, Junchi, et al.
Publicado: (2025)
Strix: Re-thinking NPU Reliability from a System Perspective
por: Guan, Jiapeng, et al.
Publicado: (2026)
por: Guan, Jiapeng, et al.
Publicado: (2026)
RecFlash: Fast Recommendation System on In-Storage Computing with Frequency-Based Data Mapping
por: Baik, Jangho, et al.
Publicado: (2026)
por: Baik, Jangho, et al.
Publicado: (2026)
Monad: Towards Cost-effective Specialization for Chiplet-based Spatial Accelerators
por: Hao, Xiaochen, et al.
Publicado: (2023)
por: Hao, Xiaochen, et al.
Publicado: (2023)
Sim-FA: A GPGPU Simulator Framework for Fine-Grained FlashAttention Pipeline Analysis
por: Zhou, Zhongchun, et al.
Publicado: (2026)
por: Zhou, Zhongchun, et al.
Publicado: (2026)
H-FA: A Hybrid Floating-Point and Logarithmic Approach to Hardware Accelerated FlashAttention
por: Alexandridis, Kosmas, et al.
Publicado: (2025)
por: Alexandridis, Kosmas, et al.
Publicado: (2025)
Nemo: A Low-Write-Amplification Cache for Tiny Objects on Log-Structured Flash Devices
por: Yang, Xufeng, et al.
Publicado: (2026)
por: Yang, Xufeng, et al.
Publicado: (2026)
FlatAttention: Dataflow and Fabric Collectives Co-Optimization for Efficient Multi-Head Attention on Tile-Based Many-PE Accelerators
por: Zhang, Chi, et al.
Publicado: (2025)
por: Zhang, Chi, et al.
Publicado: (2025)
Design of a 6-bit Threshold Inverter Quantization (TIQ) Flash Analog to Digital Converter (ADC)
por: Sarkar, Noyon Kumar, et al.
Publicado: (2025)
por: Sarkar, Noyon Kumar, et al.
Publicado: (2025)
Ultra8T: A Sub-Threshold 8T SRAM with Leakage Detection
por: Shen, Shan, et al.
Publicado: (2023)
por: Shen, Shan, et al.
Publicado: (2023)
PDA-LSTM: Knowledge-driven page data arrangement based on LSTM for LCM supression in QLC 3D NAND flash memories
por: Li, Qianhui, et al.
Publicado: (2025)
por: Li, Qianhui, et al.
Publicado: (2025)
Learning Library Cell Representations in Vector Space
por: Liang, Rongjian, et al.
Publicado: (2025)
por: Liang, Rongjian, et al.
Publicado: (2025)
Exploring and Exploiting Runtime Reconfigurable Floating Point Precision in Scientific Computing: a Case Study for Solving PDEs
por: Hao, Cong "Callie"
Publicado: (2024)
por: Hao, Cong "Callie"
Publicado: (2024)
Co-Designing Graph-based Approximate Nearest Neighbor Search at Billion Scale for Processing-in-Memory
por: Chen, Sitian, et al.
Publicado: (2026)
por: Chen, Sitian, et al.
Publicado: (2026)
Energy-Efficient QoS-Aware Scheduling for S-NUCA Many-Cores
por: Wasala, Sudam M., et al.
Publicado: (2025)
por: Wasala, Sudam M., et al.
Publicado: (2025)
Implementation and Analysis of Thermometer Encoding in DWN FPGA Accelerators
por: Mecik, Michael, et al.
Publicado: (2025)
por: Mecik, Michael, et al.
Publicado: (2025)
A Dense and Efficient Instruction Set Architecture Encoding
por: Maroun, Emad Jacob
Publicado: (2025)
por: Maroun, Emad Jacob
Publicado: (2025)
SpANNS: Optimizing Approximate Nearest Neighbor Search for Sparse Vectors Using Near Memory Processing
por: Zhang, Tianqi, et al.
Publicado: (2026)
por: Zhang, Tianqi, et al.
Publicado: (2026)
Instruction Scheduling in the Saturn Vector Unit
por: Zhao, Jerry, et al.
Publicado: (2024)
por: Zhao, Jerry, et al.
Publicado: (2024)
Register Dispersion: Reducing the Footprint of the Vector Register File in Vector Engines of Low-Cost RISC-V CPUs
por: Titopoulos, Vasileios, et al.
Publicado: (2025)
por: Titopoulos, Vasileios, et al.
Publicado: (2025)
Ejemplares similares
-
Search-in-Memory (SiM): Reliable, Versatile, and Efficient Data Matching in SSD's NAND Flash Memory Chip for Data Indexing Acceleration
por: Chen, Yun-Chih, et al.
Publicado: (2024) -
FeNOMS: Enhancing Open Modification Spectral Library Search with In-Storage Processing on Ferroelectric NAND (FeNAND) Flash
por: Pinge, Sumukh, et al.
Publicado: (2025) -
Flexible In-NAND Cryptographic Processing for Secure Flash Storage
por: Noh, Seock-Hwan, et al.
Publicado: (2025) -
HDDB: Efficient In-Storage SQL Database Search Using Hyperdimensional Computing on Ferroelectric NAND Flash
por: Zhao, Quanling, et al.
Publicado: (2025) -
AERO: Adaptive Erase Operation for Improving Lifetime and Performance of Modern NAND Flash-Based SSDs
por: Cho, Sungjun, et al.
Publicado: (2024)