ProactivePIM: Accelerating Weight-Sharing Embedding Layer with PIM for Scalable Recommendation System
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Youngsuk, Lim, Junghwan, Lee, Hyuk-Jae, Rhee, Chae Eun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
HH-PIM: Dynamic Optimization of Power and Performance with Heterogeneous-Hybrid PIM for Edge AI Devices
di: Jeon, Sangmin, et al.
Pubblicazione: (2025)
di: Jeon, Sangmin, et al.
Pubblicazione: (2025)
LP5X-PIM Sim: A High-Fidelity HW/SW Integrated Simulator for LPDDR5X-PIM
di: Cha, SangHoon, et al.
Pubblicazione: (2026)
di: Cha, SangHoon, et al.
Pubblicazione: (2026)
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-bit LLMs
di: Malekar, Jinendra, et al.
Pubblicazione: (2025)
di: Malekar, Jinendra, et al.
Pubblicazione: (2025)
PIMphony: Overcoming Bandwidth and Capacity Inefficiency in PIM-based Long-Context LLM Inference System
di: Kwon, Hyucksung, et al.
Pubblicazione: (2024)
di: Kwon, Hyucksung, et al.
Pubblicazione: (2024)
PIM-MMU: A Memory Management Unit for Accelerating Data Transfers in Commercial PIM Systems
di: Lee, Dongjae, et al.
Pubblicazione: (2024)
di: Lee, Dongjae, et al.
Pubblicazione: (2024)
Sangam: Chiplet-Based DRAM-PIM Accelerator with CXL Integration for LLM Inferencing
di: Kiyawat, Khyati, et al.
Pubblicazione: (2025)
di: Kiyawat, Khyati, et al.
Pubblicazione: (2025)
IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System
di: Seo, Minseok, et al.
Pubblicazione: (2024)
di: Seo, Minseok, et al.
Pubblicazione: (2024)
Pathfinding Future PIM Architectures by Demystifying a Commercial PIM Technology
di: Hyun, Bongjoon, et al.
Pubblicazione: (2023)
di: Hyun, Bongjoon, et al.
Pubblicazione: (2023)
PIM-malloc: A Fast and Scalable Dynamic Memory Allocator for Processing-In-Memory (PIM) Architectures
di: Lee, Dongjae, et al.
Pubblicazione: (2025)
di: Lee, Dongjae, et al.
Pubblicazione: (2025)
Inclusive-PIM: Hardware-Software Co-design for Broad Acceleration on Commercial PIM Architectures
di: Alsop, Johnathan, et al.
Pubblicazione: (2023)
di: Alsop, Johnathan, et al.
Pubblicazione: (2023)
NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing
di: Heo, Guseul, et al.
Pubblicazione: (2024)
di: Heo, Guseul, et al.
Pubblicazione: (2024)
The BRAM is the Limit: Shattering Myths, Shaping Standards, and Building Scalable PIM Accelerators
di: Kabir, MD Arafat, et al.
Pubblicazione: (2024)
di: Kabir, MD Arafat, et al.
Pubblicazione: (2024)
Dataflow-Aware PIM-Enabled Manycore Architecture for Deep Learning Workloads
di: Sharma, Harsh, et al.
Pubblicazione: (2024)
di: Sharma, Harsh, et al.
Pubblicazione: (2024)
AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization
di: Matsushima, Kosuke, et al.
Pubblicazione: (2026)
di: Matsushima, Kosuke, et al.
Pubblicazione: (2026)
Annotated PIM Bibliography
di: Kogge, Peter M.
Pubblicazione: (2026)
di: Kogge, Peter M.
Pubblicazione: (2026)
CD-PIM: A High-Bandwidth and Compute-Efficient LPDDR5-Based PIM for Low-Batch LLM Acceleration on Edge-Device
di: Lin, Ye, et al.
Pubblicazione: (2026)
di: Lin, Ye, et al.
Pubblicazione: (2026)
AME-PIM: Can Memory be Your Next Tensor Accelerator?
di: Venieri, Emanuele, et al.
Pubblicazione: (2026)
di: Venieri, Emanuele, et al.
Pubblicazione: (2026)
Sieve: Dynamic Expert-Aware PIM Acceleration for Evolving Mixture-of-Experts Models
di: Kim, Jungwoo, et al.
Pubblicazione: (2026)
di: Kim, Jungwoo, et al.
Pubblicazione: (2026)
AIM: Software and Hardware Co-design for Architecture-level IR-drop Mitigation in High-performance PIM
di: Zhang, Yuanpeng, et al.
Pubblicazione: (2025)
di: Zhang, Yuanpeng, et al.
Pubblicazione: (2025)
Membrane: Accelerating Database Analytics with Bank-Level DRAM-PIM Filtering
di: Shekar, Akhil, et al.
Pubblicazione: (2025)
di: Shekar, Akhil, et al.
Pubblicazione: (2025)
PIM-GPT: A Hybrid Process-in-Memory Accelerator for Autoregressive Transformers
di: Wu, Yuting, et al.
Pubblicazione: (2023)
di: Wu, Yuting, et al.
Pubblicazione: (2023)
Shared-PIM: Enabling Concurrent Computation and Data Flow for Faster Processing-in-DRAM
di: Mamdouh, Ahmed, et al.
Pubblicazione: (2024)
di: Mamdouh, Ahmed, et al.
Pubblicazione: (2024)
LOCALUT: Harnessing Capacity-Computation Tradeoffs for LUT-Based Inference in DRAM-PIM
di: Hong, Junguk, et al.
Pubblicazione: (2026)
di: Hong, Junguk, et al.
Pubblicazione: (2026)
DL-PIM: Improving Data Locality in Processing-in-Memory Systems
di: Tian, Parker Hao, et al.
Pubblicazione: (2025)
di: Tian, Parker Hao, et al.
Pubblicazione: (2025)
RED: Energy Optimization Framework for eDRAM-based PIM with Reconfigurable Voltage Swing and Retention-aware Scheduling
di: Kim, Jae-Young, et al.
Pubblicazione: (2025)
di: Kim, Jae-Young, et al.
Pubblicazione: (2025)
Embedded FPGA Acceleration of Brain-Like Neural Networks: Online Learning to Scalable Inference
di: Hafiz, Muhammad Ihsan Al, et al.
Pubblicazione: (2025)
di: Hafiz, Muhammad Ihsan Al, et al.
Pubblicazione: (2025)
Towards Efficient LUT-based PIM: A Scalable and Low-Power Approach for Modern Workloads
di: Khabbazan, Bahareh, et al.
Pubblicazione: (2025)
di: Khabbazan, Bahareh, et al.
Pubblicazione: (2025)
Column-wise Quantization of Weights and Partial Sums for Accurate and Efficient Compute-In-Memory Accelerators
di: Kim, Jiyoon, et al.
Pubblicazione: (2025)
di: Kim, Jiyoon, et al.
Pubblicazione: (2025)
PIM-AI: A Novel Architecture for High-Efficiency LLM Inference
di: Ortega, Cristobal, et al.
Pubblicazione: (2024)
di: Ortega, Cristobal, et al.
Pubblicazione: (2024)
LEAP: LLM Inference on Scalable PIM-NoC Architecture with Balanced Dataflow and Fine-Grained Parallelism
di: Wang, Yimin, et al.
Pubblicazione: (2025)
di: Wang, Yimin, et al.
Pubblicazione: (2025)
PIM-Opt: Demystifying Distributed Optimization Algorithms on a Real-World Processing-In-Memory System
di: Rhyner, Steve, et al.
Pubblicazione: (2024)
di: Rhyner, Steve, et al.
Pubblicazione: (2024)
Fast-OverlaPIM: A Fast Overlap-driven Mapping Framework for Processing In-Memory Neural Network Acceleration
di: Wang, Xuan, et al.
Pubblicazione: (2024)
di: Wang, Xuan, et al.
Pubblicazione: (2024)
Row-Column Hybrid Grouping for Fault-Resilient Multi-Bit Weight Representation on IMC Arrays
di: Jeon, Kang Eun, et al.
Pubblicazione: (2025)
di: Jeon, Kang Eun, et al.
Pubblicazione: (2025)
TSB: Tiny Shared Block for Efficient DNN Deployment on NVCIM Accelerators
di: Qin, Yifan, et al.
Pubblicazione: (2024)
di: Qin, Yifan, et al.
Pubblicazione: (2024)
REASON: Accelerating Probabilistic Logical Reasoning for Scalable Neuro-Symbolic Intelligence
di: Wan, Zishen, et al.
Pubblicazione: (2026)
di: Wan, Zishen, et al.
Pubblicazione: (2026)
SWAT: Scalable and Efficient Window Attention-based Transformers Acceleration on FPGAs
di: Bai, Zhenyu, et al.
Pubblicazione: (2024)
di: Bai, Zhenyu, et al.
Pubblicazione: (2024)
Scalable Processing-Near-Memory for 1M-Token LLM Inference: CXL-Enabled KV-Cache Management Beyond GPU Limits
di: Kim, Dowon, et al.
Pubblicazione: (2025)
di: Kim, Dowon, et al.
Pubblicazione: (2025)
L3: DIMM-PIM Integrated Architecture and Coordination for Scalable Long-Context LLM Inference
di: Liu, Qingyuan, et al.
Pubblicazione: (2025)
di: Liu, Qingyuan, et al.
Pubblicazione: (2025)
MoNDE: Mixture of Near-Data Experts for Large-Scale Sparse Models
di: Kim, Taehyun, et al.
Pubblicazione: (2024)
di: Kim, Taehyun, et al.
Pubblicazione: (2024)
Heterogeneous Acceleration Pipeline for Recommendation System Training
di: Adnan, Muhammad, et al.
Pubblicazione: (2022)
di: Adnan, Muhammad, et al.
Pubblicazione: (2022)
Documenti analoghi
-
HH-PIM: Dynamic Optimization of Power and Performance with Heterogeneous-Hybrid PIM for Edge AI Devices
di: Jeon, Sangmin, et al.
Pubblicazione: (2025) -
LP5X-PIM Sim: A High-Fidelity HW/SW Integrated Simulator for LPDDR5X-PIM
di: Cha, SangHoon, et al.
Pubblicazione: (2026) -
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-bit LLMs
di: Malekar, Jinendra, et al.
Pubblicazione: (2025) -
PIMphony: Overcoming Bandwidth and Capacity Inefficiency in PIM-based Long-Context LLM Inference System
di: Kwon, Hyucksung, et al.
Pubblicazione: (2024) -
PIM-MMU: A Memory Management Unit for Accelerating Data Transfers in Commercial PIM Systems
di: Lee, Dongjae, et al.
Pubblicazione: (2024)