HURRY: Highly Utilized, Reconfigurable ReRAM-based In-situ Accelerator with Multifunctionality
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shin, Hery, Kim, Jae-Young, Kim, Donghyuk, Kim, Joo-Young |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RED: Energy Optimization Framework for eDRAM-based PIM with Reconfigurable Voltage Swing and Retention-aware Scheduling
von: Kim, Jae-Young, et al.
Veröffentlicht: (2025)
von: Kim, Jae-Young, et al.
Veröffentlicht: (2025)
SAL-PIM: A Subarray-level Processing-in-Memory Architecture with LUT-based Linear Interpolation for Transformer-based Text Generation
von: Han, Wontak, et al.
Veröffentlicht: (2024)
von: Han, Wontak, et al.
Veröffentlicht: (2024)
ARAS: An Adaptive Low-Cost ReRAM-Based Accelerator for DNNs
von: Sabri, Mohammad, et al.
Veröffentlicht: (2024)
von: Sabri, Mohammad, et al.
Veröffentlicht: (2024)
Hamun: An Approximate Computation Method to Prolong the Lifespan of ReRAM-Based Accelerators
von: Sabri, Mohammad, et al.
Veröffentlicht: (2025)
von: Sabri, Mohammad, et al.
Veröffentlicht: (2025)
DIRC-RAG: Accelerating Edge RAG with Robust High-Density and High-Loading-Bandwidth Digital In-ReRAM Computation
von: Shao, Kunming, et al.
Veröffentlicht: (2025)
von: Shao, Kunming, et al.
Veröffentlicht: (2025)
Pointer: An Energy-Efficient ReRAM-based Point Cloud Recognition Accelerator with Inter-layer and Intra-layer Optimizations
von: Zhang, Qijun, et al.
Veröffentlicht: (2024)
von: Zhang, Qijun, et al.
Veröffentlicht: (2024)
Online Soft Error Tolerance in ReRAM Crossbars for Deep Learning Accelerators
von: Khezeli, Benyamin, et al.
Veröffentlicht: (2024)
von: Khezeli, Benyamin, et al.
Veröffentlicht: (2024)
All-in-Memory Stochastic Computing using ReRAM
von: de Lima, João Paulo C., et al.
Veröffentlicht: (2025)
von: de Lima, João Paulo C., et al.
Veröffentlicht: (2025)
V-Rex: Real-Time Streaming Video LLM Acceleration via Dynamic KV Cache Retrieval
von: Kim, Donghyuk, et al.
Veröffentlicht: (2025)
von: Kim, Donghyuk, et al.
Veröffentlicht: (2025)
Algorithm-hardware co-design for Energy-Efficient A/D conversion in ReRAM-based accelerators
von: Zhang, Chenguang, et al.
Veröffentlicht: (2024)
von: Zhang, Chenguang, et al.
Veröffentlicht: (2024)
Securing DRAM at Scale: ARFM-Driven Row Hammer Defense with Unveiling the Threat of Short tRC Patterns
von: Joo, Nogeun, et al.
Veröffentlicht: (2025)
von: Joo, Nogeun, et al.
Veröffentlicht: (2025)
Sensitivity-Aware Mixed-Precision Quantization for ReRAM-based Computing-in-Memory
von: Chen, Guan-Cheng, et al.
Veröffentlicht: (2025)
von: Chen, Guan-Cheng, et al.
Veröffentlicht: (2025)
FARe: Fault-Aware GNN Training on ReRAM-based PIM Accelerators
von: Dhingra, Pratyush, et al.
Veröffentlicht: (2024)
von: Dhingra, Pratyush, et al.
Veröffentlicht: (2024)
A Fully Hardware Implemented Accelerator Design in ReRAM Analog Computing without ADCs
von: Dang, Peng, et al.
Veröffentlicht: (2024)
von: Dang, Peng, et al.
Veröffentlicht: (2024)
MASQ: Accelerating Masked Diffusion via Stage-Wise Multi-Precision Quantization
von: Kim, Seeyeon, et al.
Veröffentlicht: (2026)
von: Kim, Seeyeon, et al.
Veröffentlicht: (2026)
DiSC: Resolution-Scalable Acceleration of Diffusion Models by Exploiting Sparsity and Cached Token Reuse with Hash-based Distribution
von: Yoon, Jieon, et al.
Veröffentlicht: (2026)
von: Yoon, Jieon, et al.
Veröffentlicht: (2026)
ReCross: Efficient Embedding Reduction Scheme for In-Memory Computing using ReRAM-Based Crossbar
von: Lai, Yu-Hong, et al.
Veröffentlicht: (2025)
von: Lai, Yu-Hong, et al.
Veröffentlicht: (2025)
Zero-Space Cost Fault Tolerance for Transformer-based Language Models on ReRAM
von: Li, Bingbing, et al.
Veröffentlicht: (2024)
von: Li, Bingbing, et al.
Veröffentlicht: (2024)
4T2R X-ReRAM CiM Array for Variation-tolerant, Low-power, Massively Parallel MAC Operation
von: Kihara, Fuyuki, et al.
Veröffentlicht: (2025)
von: Kihara, Fuyuki, et al.
Veröffentlicht: (2025)
ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration
von: Lee, Hangyeol, et al.
Veröffentlicht: (2026)
von: Lee, Hangyeol, et al.
Veröffentlicht: (2026)
APINT: A Full-Stack Framework for Acceleration of Privacy-Preserving Inference of Transformers based on Garbled Circuits
von: Cho, Hyunjun, et al.
Veröffentlicht: (2025)
von: Cho, Hyunjun, et al.
Veröffentlicht: (2025)
31.1 A 14.08-to-135.69Token/s ReRAM-on-Logic Stacked Outlier-Free Large-Language-Model Accelerator with Block-Clustered Weight-Compression and Adaptive Parallel-Speculative-Decoding
von: Dong, Pingcheng, et al.
Veröffentlicht: (2026)
von: Dong, Pingcheng, et al.
Veröffentlicht: (2026)
SCRec: A Scalable Computational Storage System with Statistical Sharding and Tensor-train Decomposition for Recommendation Models
von: Yang, Jinho, et al.
Veröffentlicht: (2025)
von: Yang, Jinho, et al.
Veröffentlicht: (2025)
All-in-One Analog AI Hardware: On-Chip Training and Inference with Conductive-Metal-Oxide/HfOx ReRAM Devices
von: Falcone, Donato Francesco, et al.
Veröffentlicht: (2025)
von: Falcone, Donato Francesco, et al.
Veröffentlicht: (2025)
Stuck-at Faults in ReRAM Neuromorphic Circuit Array and their Correction through Machine Learning
von: Sawal, Vedant, et al.
Veröffentlicht: (2024)
von: Sawal, Vedant, et al.
Veröffentlicht: (2024)
TL-nvSRAM-CIM: Ultra-High-Density Three-Level ReRAM-Assisted Computing-in-nvSRAM with DC-Power Free Restore and Ternary MAC Operations
von: Wang, Dengfeng, et al.
Veröffentlicht: (2023)
von: Wang, Dengfeng, et al.
Veröffentlicht: (2023)
ADOR: A Design Exploration Framework for LLM Serving with Enhanced Latency and Throughput
von: Kim, Junsoo, et al.
Veröffentlicht: (2025)
von: Kim, Junsoo, et al.
Veröffentlicht: (2025)
IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System
von: Seo, Minseok, et al.
Veröffentlicht: (2024)
von: Seo, Minseok, et al.
Veröffentlicht: (2024)
Sparse-on-Dense: Area and Energy-Efficient Computing of Sparse Neural Networks on Dense Matrix Multiplication Accelerators
von: Yoon, Hyunsung, et al.
Veröffentlicht: (2026)
von: Yoon, Hyunsung, et al.
Veröffentlicht: (2026)
LPU: A Latency-Optimized and Highly Scalable Processor for Large Language Model Inference
von: Moon, Seungjae, et al.
Veröffentlicht: (2024)
von: Moon, Seungjae, et al.
Veröffentlicht: (2024)
Oaken: Fast and Efficient LLM Serving with Online-Offline Hybrid KV Cache Quantization
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
RPCAcc: A High-Performance and Reconfigurable PCIe-attached RPC Accelerator
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
FlexNeRFer: A Multi-Dataflow, Adaptive Sparsity-Aware Accelerator for On-Device NeRF Rendering
von: Noh, Seock-Hwan, et al.
Veröffentlicht: (2025)
von: Noh, Seock-Hwan, et al.
Veröffentlicht: (2025)
Towards Performance-Aware Allocation for Accelerated Machine Learning on GPU-SSD Systems
von: Gundawar, Ayush, et al.
Veröffentlicht: (2024)
von: Gundawar, Ayush, et al.
Veröffentlicht: (2024)
Token-Picker: Accelerating Attention in Text Generation with Minimized Memory Transfer via Probability Estimation
von: Park, Junyoung, et al.
Veröffentlicht: (2024)
von: Park, Junyoung, et al.
Veröffentlicht: (2024)
Reuse Detector: Improving the Management of STT-RAM SLLCs
von: RodrÍguez-RodrÍguez, Roberto, et al.
Veröffentlicht: (2024)
von: RodrÍguez-RodrÍguez, Roberto, et al.
Veröffentlicht: (2024)
Hardware-Software Co-Design for Accelerating Transformer Inference Leveraging Compute-in-Memory
von: Kim, Dong Eun, et al.
Veröffentlicht: (2025)
von: Kim, Dong Eun, et al.
Veröffentlicht: (2025)
MTU: The Multifunction Tree Unit for Accelerating Zero-Knowledge Proofs
von: Mo, Jianqiao, et al.
Veröffentlicht: (2025)
von: Mo, Jianqiao, et al.
Veröffentlicht: (2025)
STAR: Improving Lifetime and Performance of High-Capacity Modern SSDs Using State-Aware Randomizer
von: Kwon, Omin, et al.
Veröffentlicht: (2025)
von: Kwon, Omin, et al.
Veröffentlicht: (2025)
Bandwidth-Effective DRAM Cache for GPUs with Storage-Class Memory
von: Hong, Jeongmin, et al.
Veröffentlicht: (2024)
von: Hong, Jeongmin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
RED: Energy Optimization Framework for eDRAM-based PIM with Reconfigurable Voltage Swing and Retention-aware Scheduling
von: Kim, Jae-Young, et al.
Veröffentlicht: (2025) -
SAL-PIM: A Subarray-level Processing-in-Memory Architecture with LUT-based Linear Interpolation for Transformer-based Text Generation
von: Han, Wontak, et al.
Veröffentlicht: (2024) -
ARAS: An Adaptive Low-Cost ReRAM-Based Accelerator for DNNs
von: Sabri, Mohammad, et al.
Veröffentlicht: (2024) -
Hamun: An Approximate Computation Method to Prolong the Lifespan of ReRAM-Based Accelerators
von: Sabri, Mohammad, et al.
Veröffentlicht: (2025) -
DIRC-RAG: Accelerating Edge RAG with Robust High-Density and High-Loading-Bandwidth Digital In-ReRAM Computation
von: Shao, Kunming, et al.
Veröffentlicht: (2025)