U-SWIM: Universal Selective Write-Verify for Computing-in-Memory Neural Accelerators
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yan, Zheyu, Hu, Xiaobo Sharon, Shi, Yiyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
When Small Variations Become Big Failures: Reliability Challenges in Compute-in-Memory Neural Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2026)
von: Qin, Yifan, et al.
Veröffentlicht: (2026)
Special Session: Sustainable Deployment of Deep Neural Networks on Non-Volatile Compute-in-Memory Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2025)
von: Qin, Yifan, et al.
Veröffentlicht: (2025)
NeFT: Negative Feedback Training to Improve Robustness of Compute-In-Memory DNN Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2023)
von: Qin, Yifan, et al.
Veröffentlicht: (2023)
TSB: Tiny Shared Block for Efficient DNN Deployment on NVCIM Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2024)
von: Qin, Yifan, et al.
Veröffentlicht: (2024)
CHIME: Chiplet-based Heterogeneous Near-Memory Acceleration for Edge Multimodal LLM Inference
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
Memory Is All You Need: An Overview of Compute-in-Memory Architectures for Accelerating Large Language Model Inference
von: Wolters, Christopher, et al.
Veröffentlicht: (2024)
von: Wolters, Christopher, et al.
Veröffentlicht: (2024)
Binary Weight Multi-Bit Activation Quantization for Compute-in-Memory CNN Accelerators
von: Zhou, Wenyong, et al.
Veröffentlicht: (2025)
von: Zhou, Wenyong, et al.
Veröffentlicht: (2025)
Sustainable Transformer Neural Network Acceleration with Stochastic Photonic Computing
von: Afifi, S., et al.
Veröffentlicht: (2026)
von: Afifi, S., et al.
Veröffentlicht: (2026)
Layer-wise Weight Selection for Power-Efficient Neural Network Acceleration
von: Fang, Jiaxun, et al.
Veröffentlicht: (2025)
von: Fang, Jiaxun, et al.
Veröffentlicht: (2025)
SafeCiM: Investigating Resilience of Hybrid Floating-Point Compute-in-Memory Deep Learning Accelerators
von: Bhattacharya, Swastik, et al.
Veröffentlicht: (2025)
von: Bhattacharya, Swastik, et al.
Veröffentlicht: (2025)
Optical Computing for Deep Neural Network Acceleration: Foundations, Recent Developments, and Emerging Directions
von: Pasricha, Sudeep
Veröffentlicht: (2024)
von: Pasricha, Sudeep
Veröffentlicht: (2024)
EPIM: Efficient Processing-In-Memory Accelerators based on Epitome
von: Wang, Chenyu, et al.
Veröffentlicht: (2023)
von: Wang, Chenyu, et al.
Veröffentlicht: (2023)
Efficient In-Memory Acceleration of Sparse Block Diagonal LLMs
von: de Lima, João Paulo Cardoso, et al.
Veröffentlicht: (2025)
von: de Lima, João Paulo Cardoso, et al.
Veröffentlicht: (2025)
A Computing-in-Memory-based One-Class Hyperdimensional Computing Model for Outlier Detection
von: Wang, Ruixuan, et al.
Veröffentlicht: (2023)
von: Wang, Ruixuan, et al.
Veröffentlicht: (2023)
End-to-End Transformer Acceleration Through Processing-in-Memory Architectures
von: Yang, Xiaoxuan, et al.
Veröffentlicht: (2025)
von: Yang, Xiaoxuan, et al.
Veröffentlicht: (2025)
Kernel Approximation using Analog In-Memory Computing
von: Büchel, Julian, et al.
Veröffentlicht: (2024)
von: Büchel, Julian, et al.
Veröffentlicht: (2024)
Reconfigurable Computing Challenge: Real-Time Graph Neural Networks for Online Event Selection in Big Science
von: Neu, Marc, et al.
Veröffentlicht: (2026)
von: Neu, Marc, et al.
Veröffentlicht: (2026)
Accelerating Computer Architecture Simulation through Machine Learning
von: Ali, Wajid, et al.
Veröffentlicht: (2024)
von: Ali, Wajid, et al.
Veröffentlicht: (2024)
OPIMA: Optical Processing-In-Memory for Convolutional Neural Network Acceleration
von: Sunny, Febin, et al.
Veröffentlicht: (2024)
von: Sunny, Febin, et al.
Veröffentlicht: (2024)
RACE-IT: A Reconfigurable Analog Computing Engine for In-Memory Transformer Acceleration
von: Zhao, Lei, et al.
Veröffentlicht: (2023)
von: Zhao, Lei, et al.
Veröffentlicht: (2023)
Pimba: A Processing-in-Memory Acceleration for Post-Transformer Large Language Model Serving
von: Kim, Wonung, et al.
Veröffentlicht: (2025)
von: Kim, Wonung, et al.
Veröffentlicht: (2025)
A Persistent-State Dataflow Accelerator for Memory-Bound Linear Attention Decode on FPGA
von: Gupta, Neelesh, et al.
Veröffentlicht: (2026)
von: Gupta, Neelesh, et al.
Veröffentlicht: (2026)
Accelerating Sparse Graph Neural Networks with Tensor Core Optimization
von: Wu, Ka Wai
Veröffentlicht: (2024)
von: Wu, Ka Wai
Veröffentlicht: (2024)
Token-Picker: Accelerating Attention in Text Generation with Minimized Memory Transfer via Probability Estimation
von: Park, Junyoung, et al.
Veröffentlicht: (2024)
von: Park, Junyoung, et al.
Veröffentlicht: (2024)
iEEG Seizure Detection with a Sparse Hyperdimensional Computing Accelerator
von: Cuyckens, Stef, et al.
Veröffentlicht: (2025)
von: Cuyckens, Stef, et al.
Veröffentlicht: (2025)
AnalogNAS-Bench: A NAS Benchmark for Analog In-Memory Computing
von: Bessalah, Aniss, et al.
Veröffentlicht: (2025)
von: Bessalah, Aniss, et al.
Veröffentlicht: (2025)
Accelerating LLM Inference with Flexible N:M Sparsity via A Fully Digital Compute-in-Memory Accelerator
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2025)
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2025)
Algorithmic Strategies for Sustainable Reuse of Neural Network Accelerators with Permanent Faults
von: Alama, Youssef A. Ait, et al.
Veröffentlicht: (2024)
von: Alama, Youssef A. Ait, et al.
Veröffentlicht: (2024)
PhotoGAN: Generative Adversarial Neural Network Acceleration with Silicon Photonics
von: Suresh, Tharini, et al.
Veröffentlicht: (2025)
von: Suresh, Tharini, et al.
Veröffentlicht: (2025)
An Analog and Digital Hybrid Attention Accelerator for Transformers with Charge-based In-memory Computing
von: Moradifirouzabadi, Ashkan, et al.
Veröffentlicht: (2024)
von: Moradifirouzabadi, Ashkan, et al.
Veröffentlicht: (2024)
Associative Memory Based Experience Replay for Deep Reinforcement Learning
von: Li, Mengyuan, et al.
Veröffentlicht: (2022)
von: Li, Mengyuan, et al.
Veröffentlicht: (2022)
Torch2Chip: An End-to-end Customizable Deep Neural Network Compression and Deployment Toolkit for Prototype Hardware Accelerator Design
von: Meng, Jian, et al.
Veröffentlicht: (2024)
von: Meng, Jian, et al.
Veröffentlicht: (2024)
Column-wise Quantization of Weights and Partial Sums for Accurate and Efficient Compute-In-Memory Accelerators
von: Kim, Jiyoon, et al.
Veröffentlicht: (2025)
von: Kim, Jiyoon, et al.
Veröffentlicht: (2025)
Designing Efficient LLM Accelerators for Edge Devices
von: Haris, Jude, et al.
Veröffentlicht: (2024)
von: Haris, Jude, et al.
Veröffentlicht: (2024)
Accelerating Neural Networks for Large Language Models and Graph Processing with Silicon Photonics
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
Exploring Quantization and Mapping Synergy in Hardware-Aware Deep Neural Network Accelerators
von: Klhufek, Jan, et al.
Veröffentlicht: (2024)
von: Klhufek, Jan, et al.
Veröffentlicht: (2024)
ARTEMIS: A Mixed Analog-Stochastic In-DRAM Accelerator for Transformer Neural Networks
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
ESACT: An End-to-End Sparse Accelerator for Compute-Intensive Transformers via Local Similarity
von: Liu, Hongxiang, et al.
Veröffentlicht: (2025)
von: Liu, Hongxiang, et al.
Veröffentlicht: (2025)
SpiDR: A Reconfigurable Digital Compute-in-Memory Spiking Neural Network Accelerator for Event-based Perception
von: Sharma, Deepika, et al.
Veröffentlicht: (2024)
von: Sharma, Deepika, et al.
Veröffentlicht: (2024)
VEXP: A Low-Cost RISC-V ISA Extension for Accelerated Softmax Computation in Transformers
von: Wang, Run, et al.
Veröffentlicht: (2025)
von: Wang, Run, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
When Small Variations Become Big Failures: Reliability Challenges in Compute-in-Memory Neural Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2026) -
Special Session: Sustainable Deployment of Deep Neural Networks on Non-Volatile Compute-in-Memory Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2025) -
NeFT: Negative Feedback Training to Improve Robustness of Compute-In-Memory DNN Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2023) -
TSB: Tiny Shared Block for Efficient DNN Deployment on NVCIM Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2024) -
CHIME: Chiplet-based Heterogeneous Near-Memory Acceleration for Edge Multimodal LLM Inference
von: Chen, Yanru, et al.
Veröffentlicht: (2025)