Heterogeneous Memory Design Exploration for AI Accelerators with a Gain Cell Memory Compiler
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Xinxin, Yan, Lixian, Liu, Shuhan, Upton, Luke, Cai, Zhuoqi, Tan, Yiming, Li, Shengman, Jana, Koustav, Li, Peijing, Cirimelli-Low, Jesse, Tambe, Thierry, Guthaus, Matthew, Wong, H. -S. Philip |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
OpenGCRAM: An Open-Source Gain Cell Compiler Enabling Design-Space Exploration for AI Workloads
von: Wang, Xinxin, et al.
Veröffentlicht: (2025)
von: Wang, Xinxin, et al.
Veröffentlicht: (2025)
GainSight: A Unified Framework for Data Lifetime Profiling and Heterogeneous Memory Composition
von: Li, Peijing, et al.
Veröffentlicht: (2025)
von: Li, Peijing, et al.
Veröffentlicht: (2025)
The Future of Memory: Limits and Opportunities
von: Dayo, Samuel, et al.
Veröffentlicht: (2025)
von: Dayo, Samuel, et al.
Veröffentlicht: (2025)
Towards Memory Specialization: A Case for Long-Term and Short-Term RAM
von: Li, Peijing, et al.
Veröffentlicht: (2025)
von: Li, Peijing, et al.
Veröffentlicht: (2025)
CINM (Cinnamon): A Compilation Infrastructure for Heterogeneous Compute In-Memory and Compute Near-Memory Paradigms
von: Khan, Asif Ali, et al.
Veröffentlicht: (2022)
von: Khan, Asif Ali, et al.
Veröffentlicht: (2022)
LLM-FSM: Scaling Large Language Models for Finite-State Reasoning in RTL Code Generation
von: Wu, Yuheng, et al.
Veröffentlicht: (2026)
von: Wu, Yuheng, et al.
Veröffentlicht: (2026)
PIMCOMP: An End-to-End DNN Compiler for Processing-In-Memory Accelerators
von: Sun, Xiaotian, et al.
Veröffentlicht: (2024)
von: Sun, Xiaotian, et al.
Veröffentlicht: (2024)
Be CIM or Be Memory: A Dual-mode-aware DNN Compiler for CIM Accelerators
von: Zhao, Shixin, et al.
Veröffentlicht: (2025)
von: Zhao, Shixin, et al.
Veröffentlicht: (2025)
An Open-Source Flow for Single-Phase, Edge-Triggered to Two-Phase, Non-Overlapping Clocking Conversion
von: Pedroso, Paolo, et al.
Veröffentlicht: (2026)
von: Pedroso, Paolo, et al.
Veröffentlicht: (2026)
Heterogeneous Memory Benchmarking Toolkit
von: Ghaemi, Golsana, et al.
Veröffentlicht: (2025)
von: Ghaemi, Golsana, et al.
Veröffentlicht: (2025)
The Hyperscale Lottery: How State-Space Models Have Sacrificed Edge Efficiency
von: Geens, Robin, et al.
Veröffentlicht: (2026)
von: Geens, Robin, et al.
Veröffentlicht: (2026)
GATMesh: Clock Mesh Timing Analysis using Graph Neural Networks
von: Khan, Muhammad Hadir, et al.
Veröffentlicht: (2025)
von: Khan, Muhammad Hadir, et al.
Veröffentlicht: (2025)
Compiler Testing With Relaxed Memory Models
von: Geeson, Luke, et al.
Veröffentlicht: (2023)
von: Geeson, Luke, et al.
Veröffentlicht: (2023)
LIMCA: LLM for Automating Analog In-Memory Computing Architecture Design Exploration
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025)
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025)
AccelCIM: Systematic Dataflow Exploration for SRAM Compute-in-Memory Accelerator
von: Xue, Chenhao, et al.
Veröffentlicht: (2026)
von: Xue, Chenhao, et al.
Veröffentlicht: (2026)
Hardware-based Heterogeneous Memory Management for Large Language Model Inference
von: Hwang, Soojin, et al.
Veröffentlicht: (2025)
von: Hwang, Soojin, et al.
Veröffentlicht: (2025)
Weak Memory Demands Model-based Compiler Testing
von: Geeson, Luke
Veröffentlicht: (2024)
von: Geeson, Luke
Veröffentlicht: (2024)
Cocco: Hardware-Mapping Co-Exploration towards Memory Capacity-Communication Optimization
von: Tan, Zhanhong, et al.
Veröffentlicht: (2024)
von: Tan, Zhanhong, et al.
Veröffentlicht: (2024)
SynDCIM: A Performance-Aware Digital Computing-in-Memory Compiler with Multi-Spec-Oriented Subcircuit Synthesis
von: Shao, Kunming, et al.
Veröffentlicht: (2024)
von: Shao, Kunming, et al.
Veröffentlicht: (2024)
HPIM: Heterogeneous Processing-In-Memory-based Accelerator for Large Language Models Inference
von: Duan, Cenlin, et al.
Veröffentlicht: (2025)
von: Duan, Cenlin, et al.
Veröffentlicht: (2025)
MemExplorer: Navigating the Heterogeneous Memory Design Space for Agentic Inference NPUs
von: Wu, Haoran, et al.
Veröffentlicht: (2026)
von: Wu, Haoran, et al.
Veröffentlicht: (2026)
SLTarch: Towards Scalable Point-Based Neural Rendering by Taming Workload Imbalance and Memory Irregularity
von: Li, Xingyang, et al.
Veröffentlicht: (2025)
von: Li, Xingyang, et al.
Veröffentlicht: (2025)
SCREME: A Scalable Framework for Resilient Memory Design
von: Li, Fan, et al.
Veröffentlicht: (2025)
von: Li, Fan, et al.
Veröffentlicht: (2025)
CHIME: Chiplet-based Heterogeneous Near-Memory Acceleration for Edge Multimodal LLM Inference
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
CIM-MLC: A Multi-level Compilation Stack for Computing-In-Memory Accelerators
von: Qu, Songyun, et al.
Veröffentlicht: (2024)
von: Qu, Songyun, et al.
Veröffentlicht: (2024)
Hemlet: A Heterogeneous Compute-in-Memory Chiplet Architecture for Vision Transformers with Group-Level Parallelism
von: Wang, Cong, et al.
Veröffentlicht: (2025)
von: Wang, Cong, et al.
Veröffentlicht: (2025)
The Case for Replication-Aware Memory-Error Protection in Disaggregated Memory
von: Volos, Haris
Veröffentlicht: (2023)
von: Volos, Haris
Veröffentlicht: (2023)
In-Memory ADC-Based Nonlinear Activation Quantization for Efficient In-Memory Computing
von: Dong, Shuai, et al.
Veröffentlicht: (2026)
von: Dong, Shuai, et al.
Veröffentlicht: (2026)
Trimma: Trimming Metadata Storage and Latency for Hybrid Memory Systems
von: Li, Yiwei, et al.
Veröffentlicht: (2024)
von: Li, Yiwei, et al.
Veröffentlicht: (2024)
Asynchronous Memory Access Unit: Exploiting Massive Parallelism for Far Memory Access
von: Wang, Luming, et al.
Veröffentlicht: (2024)
von: Wang, Luming, et al.
Veröffentlicht: (2024)
Reimagining Memory Access for LLM Inference: Compression-Aware Memory Controller Design
von: Xie, Rui, et al.
Veröffentlicht: (2025)
von: Xie, Rui, et al.
Veröffentlicht: (2025)
CIMPool: Scalable Neural Network Acceleration for Compute-In-Memory using Weight Pools
von: Li, Shurui, et al.
Veröffentlicht: (2025)
von: Li, Shurui, et al.
Veröffentlicht: (2025)
ICP: Exploiting Instruction Correlation for Prefetching Irregular Memory Accesses
von: Li, Mengming, et al.
Veröffentlicht: (2026)
von: Li, Mengming, et al.
Veröffentlicht: (2026)
Unicorn-CIM: Uncovering the Vulnerability and Improving the Resilience of High-Precision Compute-in-Memory
von: Li, Qiufeng, et al.
Veröffentlicht: (2025)
von: Li, Qiufeng, et al.
Veröffentlicht: (2025)
Efficient Open Modification Spectral Library Searching in High-Dimensional Space with Multi-Level-Cell Memory
von: Fan, Keming, et al.
Veröffentlicht: (2024)
von: Fan, Keming, et al.
Veröffentlicht: (2024)
Examem: Low-Overhead Memory Instrumentation for Intelligent Memory Systems
von: Poduval, Ashwin, et al.
Veröffentlicht: (2024)
von: Poduval, Ashwin, et al.
Veröffentlicht: (2024)
CXLRAMSim v1.0: System-Level Exploration of CXL Memory Expander Cards
von: Pathak, Karan, et al.
Veröffentlicht: (2026)
von: Pathak, Karan, et al.
Veröffentlicht: (2026)
PUMA: Efficient and Low-Cost Memory Allocation and Alignment Support for Processing-Using-Memory Architectures
von: Oliveira, Geraldo F., et al.
Veröffentlicht: (2024)
von: Oliveira, Geraldo F., et al.
Veröffentlicht: (2024)
PIM-malloc: A Fast and Scalable Dynamic Memory Allocator for Processing-In-Memory (PIM) Architectures
von: Lee, Dongjae, et al.
Veröffentlicht: (2025)
von: Lee, Dongjae, et al.
Veröffentlicht: (2025)
Enabling Efficient Transaction Processing on CXL-Based Memory Sharing
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
OpenGCRAM: An Open-Source Gain Cell Compiler Enabling Design-Space Exploration for AI Workloads
von: Wang, Xinxin, et al.
Veröffentlicht: (2025) -
GainSight: A Unified Framework for Data Lifetime Profiling and Heterogeneous Memory Composition
von: Li, Peijing, et al.
Veröffentlicht: (2025) -
The Future of Memory: Limits and Opportunities
von: Dayo, Samuel, et al.
Veröffentlicht: (2025) -
Towards Memory Specialization: A Case for Long-Term and Short-Term RAM
von: Li, Peijing, et al.
Veröffentlicht: (2025) -
CINM (Cinnamon): A Compilation Infrastructure for Heterogeneous Compute In-Memory and Compute Near-Memory Paradigms
von: Khan, Asif Ali, et al.
Veröffentlicht: (2022)