GEN-Graph: Heterogeneous PIM Accelerator for General Computational Patterns in Graph-based Dynamic Programming
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Chen, Yanru, Tian, Runyang, Li, Zheyu, Afarin, Mahbod, Xu, Weihong, Rosing, Tajana |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
CHIME: Chiplet-based Heterogeneous Near-Memory Acceleration for Edge Multimodal LLM Inference
par: Chen, Yanru, et autres
Publié: (2025)
par: Chen, Yanru, et autres
Publié: (2025)
RAPID-Graph: Recursive All-Pairs Shortest Paths Using Processing-in-Memory for Dynamic Programming on Graphs
par: Chen, Yanru, et autres
Publié: (2025)
par: Chen, Yanru, et autres
Publié: (2025)
PIM-FW: Hardware-Software Co-Design of All-pairs Shortest Paths in DRAM
par: Lu, Tsung-Han, et autres
Publié: (2025)
par: Lu, Tsung-Han, et autres
Publié: (2025)
Fast-OverlaPIM: A Fast Overlap-driven Mapping Framework for Processing In-Memory Neural Network Acceleration
par: Wang, Xuan, et autres
Publié: (2024)
par: Wang, Xuan, et autres
Publié: (2024)
HDDB: Efficient In-Storage SQL Database Search Using Hyperdimensional Computing on Ferroelectric NAND Flash
par: Zhao, Quanling, et autres
Publié: (2025)
par: Zhao, Quanling, et autres
Publié: (2025)
GenDRAM:Hardware-Software Co-Design of General Platform in DRAM
par: Lu, Tsung-Han, et autres
Publié: (2026)
par: Lu, Tsung-Han, et autres
Publié: (2026)
Proxima: Near-storage Acceleration for Graph-based Approximate Nearest Neighbor Search in 3D NAND
par: Xu, Weihong, et autres
Publié: (2023)
par: Xu, Weihong, et autres
Publié: (2023)
SLIM: A Heterogeneous Accelerator for Edge Inference of Sparse Large Language Model via Adaptive Thresholding
par: Xu, Weihong, et autres
Publié: (2025)
par: Xu, Weihong, et autres
Publié: (2025)
FaTRQ: Tiered Residual Quantization for LLM Vector Search in Far-Memory-Aware ANNS Systems
par: Zhang, Tianqi, et autres
Publié: (2026)
par: Zhang, Tianqi, et autres
Publié: (2026)
HAVEN: High-Bandwidth Flash Augmented Vector Engine for Large-Scale Approximate Nearest-Neighbor Search Acceleration
par: Hsu, Po-Kai, et autres
Publié: (2026)
par: Hsu, Po-Kai, et autres
Publié: (2026)
RapidOMS: FPGA-based Open Modification Spectral Library Searching with HD Computing
par: Pinge, Sumukh, et autres
Publié: (2024)
par: Pinge, Sumukh, et autres
Publié: (2024)
SpANNS: Optimizing Approximate Nearest Neighbor Search for Sparse Vectors Using Near Memory Processing
par: Zhang, Tianqi, et autres
Publié: (2026)
par: Zhang, Tianqi, et autres
Publié: (2026)
FSL-HDnn: A 5.7 TOPS/W End-to-end Few-shot Learning Classifier Accelerator with Feature Extraction and Hyperdimensional Computing
par: Yang, Haichao, et autres
Publié: (2024)
par: Yang, Haichao, et autres
Publié: (2024)
SpecPCM: A Low-power PCM-based In-Memory Computing Accelerator for Full-stack Mass Spectrometry Analysis
par: Fan, Keming, et autres
Publié: (2024)
par: Fan, Keming, et autres
Publié: (2024)
NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing
par: Heo, Guseul, et autres
Publié: (2024)
par: Heo, Guseul, et autres
Publié: (2024)
FSL-HDnn: A 40 nm Few-shot On-Device Learning Accelerator with Integrated Feature Extraction and Hyperdimensional Computing
par: Xu, Weihong, et autres
Publié: (2025)
par: Xu, Weihong, et autres
Publié: (2025)
Hybrid SLC-MLC RRAM Mixed-Signal Processing-in-Memory Architecture for Transformer Acceleration via Gradient Redistribution
par: Song, Chang Eun, et autres
Publié: (2025)
par: Song, Chang Eun, et autres
Publié: (2025)
Leveraging Recurrent Patterns in Graph Accelerators
par: Rahimi, Masoud, et autres
Publié: (2025)
par: Rahimi, Masoud, et autres
Publié: (2025)
CD-PIM: A High-Bandwidth and Compute-Efficient LPDDR5-Based PIM for Low-Batch LLM Acceleration on Edge-Device
par: Lin, Ye, et autres
Publié: (2026)
par: Lin, Ye, et autres
Publié: (2026)
Inclusive-PIM: Hardware-Software Co-design for Broad Acceleration on Commercial PIM Architectures
par: Alsop, Johnathan, et autres
Publié: (2023)
par: Alsop, Johnathan, et autres
Publié: (2023)
PIM-MMU: A Memory Management Unit for Accelerating Data Transfers in Commercial PIM Systems
par: Lee, Dongjae, et autres
Publié: (2024)
par: Lee, Dongjae, et autres
Publié: (2024)
IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System
par: Seo, Minseok, et autres
Publié: (2024)
par: Seo, Minseok, et autres
Publié: (2024)
Sieve: Dynamic Expert-Aware PIM Acceleration for Evolving Mixture-of-Experts Models
par: Kim, Jungwoo, et autres
Publié: (2026)
par: Kim, Jungwoo, et autres
Publié: (2026)
HH-PIM: Dynamic Optimization of Power and Performance with Heterogeneous-Hybrid PIM for Edge AI Devices
par: Jeon, Sangmin, et autres
Publié: (2025)
par: Jeon, Sangmin, et autres
Publié: (2025)
Efficient Open Modification Spectral Library Searching in High-Dimensional Space with Multi-Level-Cell Memory
par: Fan, Keming, et autres
Publié: (2024)
par: Fan, Keming, et autres
Publié: (2024)
GDR-HGNN: A Heterogeneous Graph Neural Networks Accelerator Frontend with Graph Decoupling and Recoupling
par: Xue, Runzhen, et autres
Publié: (2024)
par: Xue, Runzhen, et autres
Publié: (2024)
PIM-malloc: A Fast and Scalable Dynamic Memory Allocator for Processing-In-Memory (PIM) Architectures
par: Lee, Dongjae, et autres
Publié: (2025)
par: Lee, Dongjae, et autres
Publié: (2025)
AME-PIM: Can Memory be Your Next Tensor Accelerator?
par: Venieri, Emanuele, et autres
Publié: (2026)
par: Venieri, Emanuele, et autres
Publié: (2026)
Clo-HDnn: A 4.66 TFLOPS/W and 3.78 TOPS/W Continual On-Device Learning Accelerator with Energy-efficient Hyperdimensional Computing via Progressive Search
par: Song, Chang Eun, et autres
Publié: (2025)
par: Song, Chang Eun, et autres
Publié: (2025)
Membrane: Accelerating Database Analytics with Bank-Level DRAM-PIM Filtering
par: Shekar, Akhil, et autres
Publié: (2025)
par: Shekar, Akhil, et autres
Publié: (2025)
PIM-GPT: A Hybrid Process-in-Memory Accelerator for Autoregressive Transformers
par: Wu, Yuting, et autres
Publié: (2023)
par: Wu, Yuting, et autres
Publié: (2023)
Characterization of Real Communication Patterns and Congestion Dynamics in HPC Interconnection Networks
par: de La Rosa, Miguel Sánchez, et autres
Publié: (2026)
par: de La Rosa, Miguel Sánchez, et autres
Publié: (2026)
The BRAM is the Limit: Shattering Myths, Shaping Standards, and Building Scalable PIM Accelerators
par: Kabir, MD Arafat, et autres
Publié: (2024)
par: Kabir, MD Arafat, et autres
Publié: (2024)
Pathfinding Future PIM Architectures by Demystifying a Commercial PIM Technology
par: Hyun, Bongjoon, et autres
Publié: (2023)
par: Hyun, Bongjoon, et autres
Publié: (2023)
ProactivePIM: Accelerating Weight-Sharing Embedding Layer with PIM for Scalable Recommendation System
par: Kim, Youngsuk, et autres
Publié: (2024)
par: Kim, Youngsuk, et autres
Publié: (2024)
DL-PIM: Improving Data Locality in Processing-in-Memory Systems
par: Tian, Parker Hao, et autres
Publié: (2025)
par: Tian, Parker Hao, et autres
Publié: (2025)
Annotated PIM Bibliography
par: Kogge, Peter M.
Publié: (2026)
par: Kogge, Peter M.
Publié: (2026)
Extracting TCPIP Headers at High Speed for the Anonymized Network Traffic Graph Challenge
par: Han, Zhaoyang, et autres
Publié: (2024)
par: Han, Zhaoyang, et autres
Publié: (2024)
Graphitron: A Domain Specific Language for FPGA-based Graph Processing Accelerator Generation
par: Zhang, Xinmiao, et autres
Publié: (2024)
par: Zhang, Xinmiao, et autres
Publié: (2024)
Special Session: Sustainable Deployment of Deep Neural Networks on Non-Volatile Compute-in-Memory Accelerators
par: Qin, Yifan, et autres
Publié: (2025)
par: Qin, Yifan, et autres
Publié: (2025)
Documents similaires
-
CHIME: Chiplet-based Heterogeneous Near-Memory Acceleration for Edge Multimodal LLM Inference
par: Chen, Yanru, et autres
Publié: (2025) -
RAPID-Graph: Recursive All-Pairs Shortest Paths Using Processing-in-Memory for Dynamic Programming on Graphs
par: Chen, Yanru, et autres
Publié: (2025) -
PIM-FW: Hardware-Software Co-Design of All-pairs Shortest Paths in DRAM
par: Lu, Tsung-Han, et autres
Publié: (2025) -
Fast-OverlaPIM: A Fast Overlap-driven Mapping Framework for Processing In-Memory Neural Network Acceleration
par: Wang, Xuan, et autres
Publié: (2024) -
HDDB: Efficient In-Storage SQL Database Search Using Hyperdimensional Computing on Ferroelectric NAND Flash
par: Zhao, Quanling, et autres
Publié: (2025)