HePGA: A Heterogeneous Processing-in-Memory based GNN Training Accelerator
Fuente:
arXiv
Salvato in:
| Autori principali: | Ogbogu, Chukwufumnanya, Narang, Gaurav, Joardar, Biresh Kumar, Doppa, Janardhan Rao, Chakrabarty, Krishnendu, Pande, Partha Pratim |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FARe: Fault-Aware GNN Training on ReRAM-based PIM Accelerators
di: Dhingra, Pratyush, et al.
Pubblicazione: (2024)
di: Dhingra, Pratyush, et al.
Pubblicazione: (2024)
Dataflow-Aware PIM-Enabled Manycore Architecture for Deep Learning Workloads
di: Sharma, Harsh, et al.
Pubblicazione: (2024)
di: Sharma, Harsh, et al.
Pubblicazione: (2024)
HeTraX: Energy Efficient 3D Heterogeneous Manycore Architecture for Transformer Acceleration
di: Dhingra, Pratyush, et al.
Pubblicazione: (2024)
di: Dhingra, Pratyush, et al.
Pubblicazione: (2024)
Look-Up Table based Neural Network Hardware
di: Sen, Ovishake, et al.
Pubblicazione: (2024)
di: Sen, Ovishake, et al.
Pubblicazione: (2024)
Atleus: Accelerating Transformers on the Edge Enabled by 3D Heterogeneous Manycore Architectures
di: Dhingra, Pratyush, et al.
Pubblicazione: (2025)
di: Dhingra, Pratyush, et al.
Pubblicazione: (2025)
Multi-Objective Optimization of ReRAM Crossbars for Robust DNN Inferencing under Stochastic Noise
di: Yang, Xiaoxuan, et al.
Pubblicazione: (2021)
di: Yang, Xiaoxuan, et al.
Pubblicazione: (2021)
A Heterogeneous Chiplet Architecture for Accelerating End-to-End Transformer Models
di: Sharma, Harsh, et al.
Pubblicazione: (2023)
di: Sharma, Harsh, et al.
Pubblicazione: (2023)
Designing High-Performance and Thermally Feasible Multi-Chiplet Architectures enabled by Non-bendable Glass Interposer
di: Sharma, Harsh, et al.
Pubblicazione: (2025)
di: Sharma, Harsh, et al.
Pubblicazione: (2025)
THERMOS: Thermally-Aware Multi-Objective Scheduling of AI Workloads on Heterogeneous Multi-Chiplet PIM Architectures
di: Kanani, Alish, et al.
Pubblicazione: (2025)
di: Kanani, Alish, et al.
Pubblicazione: (2025)
H3PIMAP: A Heterogeneity-Aware Multi-Objective DNN Mapping Framework on Electronic-Photonic Processing-in-Memory Architectures
di: Yin, Ziang, et al.
Pubblicazione: (2025)
di: Yin, Ziang, et al.
Pubblicazione: (2025)
CHIPSIM: A Co-Simulation Framework for Deep Learning on Chiplet-Based Systems
di: Pfromm, Lukas, et al.
Pubblicazione: (2025)
di: Pfromm, Lukas, et al.
Pubblicazione: (2025)
Attacking AI Accelerators by Leveraging Arithmetic Properties of Addition
di: Heidary, Masoud, et al.
Pubblicazione: (2026)
di: Heidary, Masoud, et al.
Pubblicazione: (2026)
Building Reliable Arithmetic Multipliers Under NBTI Aging and Process Variations
di: Heidary, Masoud, et al.
Pubblicazione: (2026)
di: Heidary, Masoud, et al.
Pubblicazione: (2026)
STAMP-2.5D: Structural and Thermal Aware Methodology for Placement in 2.5D Integration
di: Parekh, Varun Darshana, et al.
Pubblicazione: (2025)
di: Parekh, Varun Darshana, et al.
Pubblicazione: (2025)
OPIMA: Optical Processing-In-Memory for Convolutional Neural Network Acceleration
di: Sunny, Febin, et al.
Pubblicazione: (2024)
di: Sunny, Febin, et al.
Pubblicazione: (2024)
Modeling and Simulation Frameworks for Processing-in-Memory Architectures
di: Aghaei, Mahdi, et al.
Pubblicazione: (2025)
di: Aghaei, Mahdi, et al.
Pubblicazione: (2025)
MFIT: Multi-Fidelity Thermal Modeling for 2.5D and 3D Multi-Chiplet Architectures
di: Pfromm, Lukas, et al.
Pubblicazione: (2024)
di: Pfromm, Lukas, et al.
Pubblicazione: (2024)
XL-HD: Extended Learning in Hyperdimensional Computing via Deterministic Projections for In-Memory Accelerators
di: Moon, Sabrina Hassan, et al.
Pubblicazione: (2026)
di: Moon, Sabrina Hassan, et al.
Pubblicazione: (2026)
StoX-Net: Stochastic Processing of Partial Sums for Efficient In-Memory Computing DNN Accelerators
di: Rogers, Ethan G, et al.
Pubblicazione: (2024)
di: Rogers, Ethan G, et al.
Pubblicazione: (2024)
Architectural and System Implications of CXL-enabled Tiered Memory
di: Yang, Yujie, et al.
Pubblicazione: (2025)
di: Yang, Yujie, et al.
Pubblicazione: (2025)
MPAI: A Co-Processing Architecture with MPSoC & AI Accelerators for Vision Applications in Space
di: Leon, Vasileios, et al.
Pubblicazione: (2024)
di: Leon, Vasileios, et al.
Pubblicazione: (2024)
A 5T-2MTJ STT-assisted Spin Orbit Torque based Ternary Content Addressable Memory for Hardware Accelerators
di: Narla, Siri, et al.
Pubblicazione: (2024)
di: Narla, Siri, et al.
Pubblicazione: (2024)
A Survey on Deep Learning Hardware Accelerators for Heterogeneous HPC Platforms
di: Silvano, Cristina, et al.
Pubblicazione: (2023)
di: Silvano, Cristina, et al.
Pubblicazione: (2023)
Hardware Accelerators for Artificial Intelligence
di: Ahsan, S M Mojahidul, et al.
Pubblicazione: (2024)
di: Ahsan, S M Mojahidul, et al.
Pubblicazione: (2024)
RACE-IT: A Reconfigurable Analog Computing Engine for In-Memory Transformer Acceleration
di: Zhao, Lei, et al.
Pubblicazione: (2023)
di: Zhao, Lei, et al.
Pubblicazione: (2023)
Accelerating Hybrid XOR$-$CNF Boolean Satisfiability Problems Natively with In-Memory Computing
di: Im, Haesol, et al.
Pubblicazione: (2025)
di: Im, Haesol, et al.
Pubblicazione: (2025)
Computing High-Degree Polynomial Gradients in Memory
di: Bhattacharya, T., et al.
Pubblicazione: (2024)
di: Bhattacharya, T., et al.
Pubblicazione: (2024)
A Precision-Optimized Fixed-Point Near-Memory Digital Processing Unit for Analog In-Memory Computing
di: Ferro, Elena, et al.
Pubblicazione: (2024)
di: Ferro, Elena, et al.
Pubblicazione: (2024)
Hybrid Temporal Computing for Lower Power Hardware Accelerators
di: Tasnim, Maliha, et al.
Pubblicazione: (2024)
di: Tasnim, Maliha, et al.
Pubblicazione: (2024)
FPIA: Field-Programmable Ising Arrays with In-Memory Computing
di: Hutchinson, George Higgins, et al.
Pubblicazione: (2024)
di: Hutchinson, George Higgins, et al.
Pubblicazione: (2024)
All-in-Memory Stochastic Computing using ReRAM
di: de Lima, João Paulo C., et al.
Pubblicazione: (2025)
di: de Lima, João Paulo C., et al.
Pubblicazione: (2025)
RETENTION: Resource-Efficient Tree-Based Ensemble Model Acceleration with Content-Addressable Memory
di: Liao, Yi-Chun, et al.
Pubblicazione: (2025)
di: Liao, Yi-Chun, et al.
Pubblicazione: (2025)
Architecture-Level Modeling of Photonic Deep Neural Network Accelerators
di: Andrulis, Tanner, et al.
Pubblicazione: (2024)
di: Andrulis, Tanner, et al.
Pubblicazione: (2024)
Apparate: Evading Memory Hierarchy with GodSpeed Wireless-on-Chip
di: GS, Nitesh Narayana, et al.
Pubblicazione: (2024)
di: GS, Nitesh Narayana, et al.
Pubblicazione: (2024)
MASIM: An Efficient Multi-Array Scheduler for In-Memory SIMD Computation
di: Qian, Xingyue, et al.
Pubblicazione: (2024)
di: Qian, Xingyue, et al.
Pubblicazione: (2024)
Count2Multiply: Reliable In-Memory High-Radix Counting
di: de Lima, João Paulo Cardoso, et al.
Pubblicazione: (2024)
di: de Lima, João Paulo Cardoso, et al.
Pubblicazione: (2024)
Antiferromagnetic Tunnel Junctions (AFMTJs) for In-Memory Computing: Modeling and Case Study
di: Choudhary, Yousuf, et al.
Pubblicazione: (2026)
di: Choudhary, Yousuf, et al.
Pubblicazione: (2026)
Bayes2IMC: In-Memory Computing for Bayesian Binary Neural Networks
di: Katti, Prabodh, et al.
Pubblicazione: (2024)
di: Katti, Prabodh, et al.
Pubblicazione: (2024)
Content Addressable Memory Design with Reference Resistor for Improved Search Resolution
di: Narla, Siri, et al.
Pubblicazione: (2025)
di: Narla, Siri, et al.
Pubblicazione: (2025)
Online Soft Error Tolerance in ReRAM Crossbars for Deep Learning Accelerators
di: Khezeli, Benyamin, et al.
Pubblicazione: (2024)
di: Khezeli, Benyamin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
FARe: Fault-Aware GNN Training on ReRAM-based PIM Accelerators
di: Dhingra, Pratyush, et al.
Pubblicazione: (2024) -
Dataflow-Aware PIM-Enabled Manycore Architecture for Deep Learning Workloads
di: Sharma, Harsh, et al.
Pubblicazione: (2024) -
HeTraX: Energy Efficient 3D Heterogeneous Manycore Architecture for Transformer Acceleration
di: Dhingra, Pratyush, et al.
Pubblicazione: (2024) -
Look-Up Table based Neural Network Hardware
di: Sen, Ovishake, et al.
Pubblicazione: (2024) -
Atleus: Accelerating Transformers on the Edge Enabled by 3D Heterogeneous Manycore Architectures
di: Dhingra, Pratyush, et al.
Pubblicazione: (2025)