PIM-MMU: A Memory Management Unit for Accelerating Data Transfers in Commercial PIM Systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Dongjae, Hyun, Bongjoon, Kim, Taehun, Rhu, Minsoo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Pathfinding Future PIM Architectures by Demystifying a Commercial PIM Technology
di: Hyun, Bongjoon, et al.
Pubblicazione: (2023)
di: Hyun, Bongjoon, et al.
Pubblicazione: (2023)
PIM-malloc: A Fast and Scalable Dynamic Memory Allocator for Processing-In-Memory (PIM) Architectures
di: Lee, Dongjae, et al.
Pubblicazione: (2025)
di: Lee, Dongjae, et al.
Pubblicazione: (2025)
Inclusive-PIM: Hardware-Software Co-design for Broad Acceleration on Commercial PIM Architectures
di: Alsop, Johnathan, et al.
Pubblicazione: (2023)
di: Alsop, Johnathan, et al.
Pubblicazione: (2023)
IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System
di: Seo, Minseok, et al.
Pubblicazione: (2024)
di: Seo, Minseok, et al.
Pubblicazione: (2024)
Mamba-X: An End-to-End Vision Mamba Accelerator for Edge Computing Devices
di: Yoon, Dongho, et al.
Pubblicazione: (2025)
di: Yoon, Dongho, et al.
Pubblicazione: (2025)
ProactivePIM: Accelerating Weight-Sharing Embedding Layer with PIM for Scalable Recommendation System
di: Kim, Youngsuk, et al.
Pubblicazione: (2024)
di: Kim, Youngsuk, et al.
Pubblicazione: (2024)
DL-PIM: Improving Data Locality in Processing-in-Memory Systems
di: Tian, Parker Hao, et al.
Pubblicazione: (2025)
di: Tian, Parker Hao, et al.
Pubblicazione: (2025)
AME-PIM: Can Memory be Your Next Tensor Accelerator?
di: Venieri, Emanuele, et al.
Pubblicazione: (2026)
di: Venieri, Emanuele, et al.
Pubblicazione: (2026)
PIM-GPT: A Hybrid Process-in-Memory Accelerator for Autoregressive Transformers
di: Wu, Yuting, et al.
Pubblicazione: (2023)
di: Wu, Yuting, et al.
Pubblicazione: (2023)
NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing
di: Heo, Guseul, et al.
Pubblicazione: (2024)
di: Heo, Guseul, et al.
Pubblicazione: (2024)
Annotated PIM Bibliography
di: Kogge, Peter M.
Pubblicazione: (2026)
di: Kogge, Peter M.
Pubblicazione: (2026)
PreSto: An In-Storage Data Preprocessing System for Training Recommendation Models
di: Lee, Yunjae, et al.
Pubblicazione: (2024)
di: Lee, Yunjae, et al.
Pubblicazione: (2024)
CD-PIM: A High-Bandwidth and Compute-Efficient LPDDR5-Based PIM for Low-Batch LLM Acceleration on Edge-Device
di: Lin, Ye, et al.
Pubblicazione: (2026)
di: Lin, Ye, et al.
Pubblicazione: (2026)
HH-PIM: Dynamic Optimization of Power and Performance with Heterogeneous-Hybrid PIM for Edge AI Devices
di: Jeon, Sangmin, et al.
Pubblicazione: (2025)
di: Jeon, Sangmin, et al.
Pubblicazione: (2025)
Sieve: Dynamic Expert-Aware PIM Acceleration for Evolving Mixture-of-Experts Models
di: Kim, Jungwoo, et al.
Pubblicazione: (2026)
di: Kim, Jungwoo, et al.
Pubblicazione: (2026)
Membrane: Accelerating Database Analytics with Bank-Level DRAM-PIM Filtering
di: Shekar, Akhil, et al.
Pubblicazione: (2025)
di: Shekar, Akhil, et al.
Pubblicazione: (2025)
Fast-OverlaPIM: A Fast Overlap-driven Mapping Framework for Processing In-Memory Neural Network Acceleration
di: Wang, Xuan, et al.
Pubblicazione: (2024)
di: Wang, Xuan, et al.
Pubblicazione: (2024)
PIMfused: Near-Bank DRAM-PIM with Fused-layer Dataflow for CNN Data Transfer Optimization
di: Yang, Simei, et al.
Pubblicazione: (2025)
di: Yang, Simei, et al.
Pubblicazione: (2025)
LP5X-PIM Sim: A High-Fidelity HW/SW Integrated Simulator for LPDDR5X-PIM
di: Cha, SangHoon, et al.
Pubblicazione: (2026)
di: Cha, SangHoon, et al.
Pubblicazione: (2026)
LOCALUT: Harnessing Capacity-Computation Tradeoffs for LUT-Based Inference in DRAM-PIM
di: Hong, Junguk, et al.
Pubblicazione: (2026)
di: Hong, Junguk, et al.
Pubblicazione: (2026)
The BRAM is the Limit: Shattering Myths, Shaping Standards, and Building Scalable PIM Accelerators
di: Kabir, MD Arafat, et al.
Pubblicazione: (2024)
di: Kabir, MD Arafat, et al.
Pubblicazione: (2024)
A$^3$PIM: An Automated, Analytic and Accurate Processing-in-Memory Offloader
di: Jiang, Qingcai, et al.
Pubblicazione: (2024)
di: Jiang, Qingcai, et al.
Pubblicazione: (2024)
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-bit LLMs
di: Malekar, Jinendra, et al.
Pubblicazione: (2025)
di: Malekar, Jinendra, et al.
Pubblicazione: (2025)
PyPIM: Integrating Digital Processing-in-Memory from Microarchitectural Design to Python Tensors
di: Leitersdorf, Orian, et al.
Pubblicazione: (2023)
di: Leitersdorf, Orian, et al.
Pubblicazione: (2023)
SAL-PIM: A Subarray-level Processing-in-Memory Architecture with LUT-based Linear Interpolation for Transformer-based Text Generation
di: Han, Wontak, et al.
Pubblicazione: (2024)
di: Han, Wontak, et al.
Pubblicazione: (2024)
Shared-PIM: Enabling Concurrent Computation and Data Flow for Faster Processing-in-DRAM
di: Mamdouh, Ahmed, et al.
Pubblicazione: (2024)
di: Mamdouh, Ahmed, et al.
Pubblicazione: (2024)
SwarmIO: Towards 100 Million IOPS SSD Emulation for Next-generation GPU-centric Storage Systems
di: Kim, Hyeseong, et al.
Pubblicazione: (2026)
di: Kim, Hyeseong, et al.
Pubblicazione: (2026)
Toleo: Scaling Freshness to Tera-scale Memory using CXL and PIM
di: Dong, Juechu, et al.
Pubblicazione: (2024)
di: Dong, Juechu, et al.
Pubblicazione: (2024)
PIMphony: Overcoming Bandwidth and Capacity Inefficiency in PIM-based Long-Context LLM Inference System
di: Kwon, Hyucksung, et al.
Pubblicazione: (2024)
di: Kwon, Hyucksung, et al.
Pubblicazione: (2024)
Sangam: Chiplet-Based DRAM-PIM Accelerator with CXL Integration for LLM Inferencing
di: Kiyawat, Khyati, et al.
Pubblicazione: (2025)
di: Kiyawat, Khyati, et al.
Pubblicazione: (2025)
RED: Energy Optimization Framework for eDRAM-based PIM with Reconfigurable Voltage Swing and Retention-aware Scheduling
di: Kim, Jae-Young, et al.
Pubblicazione: (2025)
di: Kim, Jae-Young, et al.
Pubblicazione: (2025)
The Cost of Dynamic Reasoning: Demystifying AI Agents and Test-Time Scaling from an AI Infrastructure Perspective
di: Kim, Jiin, et al.
Pubblicazione: (2025)
di: Kim, Jiin, et al.
Pubblicazione: (2025)
PIM Is All You Need: A CXL-Enabled GPU-Free System for Large Language Model Inference
di: Gu, Yufeng, et al.
Pubblicazione: (2025)
di: Gu, Yufeng, et al.
Pubblicazione: (2025)
JSPIM: A Skew-Aware PIM Accelerator for High-Performance Databases Join and Select Operations
di: Tajdari, Sabiha, et al.
Pubblicazione: (2025)
di: Tajdari, Sabiha, et al.
Pubblicazione: (2025)
UpANNS: Enhancing Billion-Scale ANNS Efficiency with Real-World PIM Architecture
di: Chen, Sitian, et al.
Pubblicazione: (2024)
di: Chen, Sitian, et al.
Pubblicazione: (2024)
Towards Efficient SRAM-PIM Architecture Design by Exploiting Unstructured Bit-Level Sparsity
di: Duan, Cenlin, et al.
Pubblicazione: (2024)
di: Duan, Cenlin, et al.
Pubblicazione: (2024)
PIM-FW: Hardware-Software Co-Design of All-pairs Shortest Paths in DRAM
di: Lu, Tsung-Han, et al.
Pubblicazione: (2025)
di: Lu, Tsung-Han, et al.
Pubblicazione: (2025)
Dissecting and Re-architecting 3D NAND Flash PIM Arrays for Efficient Single-Batch Token Generation in LLMs
di: Jang, Yongjoo, et al.
Pubblicazione: (2025)
di: Jang, Yongjoo, et al.
Pubblicazione: (2025)
Towards Efficient LUT-based PIM: A Scalable and Low-Power Approach for Modern Workloads
di: Khabbazan, Bahareh, et al.
Pubblicazione: (2025)
di: Khabbazan, Bahareh, et al.
Pubblicazione: (2025)
AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization
di: Matsushima, Kosuke, et al.
Pubblicazione: (2026)
di: Matsushima, Kosuke, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Pathfinding Future PIM Architectures by Demystifying a Commercial PIM Technology
di: Hyun, Bongjoon, et al.
Pubblicazione: (2023) -
PIM-malloc: A Fast and Scalable Dynamic Memory Allocator for Processing-In-Memory (PIM) Architectures
di: Lee, Dongjae, et al.
Pubblicazione: (2025) -
Inclusive-PIM: Hardware-Software Co-design for Broad Acceleration on Commercial PIM Architectures
di: Alsop, Johnathan, et al.
Pubblicazione: (2023) -
IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System
di: Seo, Minseok, et al.
Pubblicazione: (2024) -
Mamba-X: An End-to-End Vision Mamba Accelerator for Edge Computing Devices
di: Yoon, Dongho, et al.
Pubblicazione: (2025)