Learnable Sparsification of Die-to-Die Communication via Spike-Based Encoding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nardone, Joshua, Zhu, Ruijie, Callenes, Joseph, Elbtity, Mohammed E., Zand, Ramtin, Eshraghian, Jason |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Flex-TPU: A Flexible TPU with Runtime Reconfigurable Dataflow Architecture
von: Elbtity, Mohammed, et al.
Veröffentlicht: (2024)
von: Elbtity, Mohammed, et al.
Veröffentlicht: (2024)
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-bit LLMs
von: Malekar, Jinendra, et al.
Veröffentlicht: (2025)
von: Malekar, Jinendra, et al.
Veröffentlicht: (2025)
TPU-Gen: LLM-Driven Custom Tensor Processing Unit Generator
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025)
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025)
Analytical Heterogeneous Die-to-Die 3D Placement with Macros
von: Zhao, Yuxuan, et al.
Veröffentlicht: (2024)
von: Zhao, Yuxuan, et al.
Veröffentlicht: (2024)
CrossNAS: A Cross-Layer Neural Architecture Search Framework for PIM Systems
von: Amin, Md Hasibul, et al.
Veröffentlicht: (2025)
von: Amin, Md Hasibul, et al.
Veröffentlicht: (2025)
LIMCA: LLM for Automating Analog In-Memory Computing Architecture Design Exploration
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025)
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025)
Interconnect-Aware Logic Resynthesis for Multi-Die FPGAs
von: Wang, Xiaoke, et al.
Veröffentlicht: (2026)
von: Wang, Xiaoke, et al.
Veröffentlicht: (2026)
FedChip: Federated LLM for Artificial Intelligence Accelerator Chip Design
von: Nazzal, Mahmoud, et al.
Veröffentlicht: (2025)
von: Nazzal, Mahmoud, et al.
Veröffentlicht: (2025)
Fleet: Hierarchical Task-based Abstraction for Megakernels on Multi-Die GPUs
von: Chowdhary, Sangeeta, et al.
Veröffentlicht: (2026)
von: Chowdhary, Sangeeta, et al.
Veröffentlicht: (2026)
LEAPS: Topological-Layout-Adaptable Multi-Die FPGA Placement for Super Long Line Minimization
von: Di, Zhixiong, et al.
Veröffentlicht: (2023)
von: Di, Zhixiong, et al.
Veröffentlicht: (2023)
Lightator: An Optical Near-Sensor Accelerator with Compressive Acquisition Enabling Versatile Image Processing
von: Morsali, Mehrdad, et al.
Veröffentlicht: (2024)
von: Morsali, Mehrdad, et al.
Veröffentlicht: (2024)
To Spike or Not To Spike: A Digital Hardware Perspective on Deep Learning Acceleration
von: Ottati, Fabrizio, et al.
Veröffentlicht: (2023)
von: Ottati, Fabrizio, et al.
Veröffentlicht: (2023)
DreamRAM: A Fine-Grained Configurable Design Space Modeling Tool for Custom 3D Die-Stacked DRAM
von: Cai, Victor, et al.
Veröffentlicht: (2025)
von: Cai, Victor, et al.
Veröffentlicht: (2025)
Prosperity: Accelerating Spiking Neural Networks via Product Sparsity
von: Wei, Chiyue, et al.
Veröffentlicht: (2025)
von: Wei, Chiyue, et al.
Veröffentlicht: (2025)
SpikeStream: Accelerating Spiking Neural Network Inference on RISC-V Clusters with Sparse Computation Extensions
von: Manoni, Simone, et al.
Veröffentlicht: (2025)
von: Manoni, Simone, et al.
Veröffentlicht: (2025)
Optimizing Neural Networks with Learnable Non-Linear Activation Functions via Lookup-Based FPGA Acceleration
von: Yin, Mengyuan, et al.
Veröffentlicht: (2025)
von: Yin, Mengyuan, et al.
Veröffentlicht: (2025)
An Efficient Sparse Hardware Accelerator for Spike-Driven Transformer
von: Li, Zhengke, et al.
Veröffentlicht: (2025)
von: Li, Zhengke, et al.
Veröffentlicht: (2025)
ATLAAS: Automatic Tensor-Level Abstraction of Accelerator Semantics
von: Gao, Ruijie, et al.
Veröffentlicht: (2026)
von: Gao, Ruijie, et al.
Veröffentlicht: (2026)
Xpikeformer: Hybrid Analog-Digital Hardware Acceleration for Spiking Transformers
von: Song, Zihang, et al.
Veröffentlicht: (2024)
von: Song, Zihang, et al.
Veröffentlicht: (2024)
A PVT-Resilient Subthreshold SRAM-Based In-Memory Computing Accelerator with In-Situ Regulation for Energy-Efficient Spiking Neural Networks
von: Kao, Shih-Hang, et al.
Veröffentlicht: (2026)
von: Kao, Shih-Hang, et al.
Veröffentlicht: (2026)
Implementation and Analysis of Thermometer Encoding in DWN FPGA Accelerators
von: Mecik, Michael, et al.
Veröffentlicht: (2025)
von: Mecik, Michael, et al.
Veröffentlicht: (2025)
A Dense and Efficient Instruction Set Architecture Encoding
von: Maroun, Emad Jacob
Veröffentlicht: (2025)
von: Maroun, Emad Jacob
Veröffentlicht: (2025)
An Event-Driven Spiking Compute-In-Memory Macro based on SOT-MRAM
von: Yu, Deyang, et al.
Veröffentlicht: (2025)
von: Yu, Deyang, et al.
Veröffentlicht: (2025)
Hardware Efficient Accelerator for Spiking Transformer With Reconfigurable Parallel Time Step Computing
von: Chen, Bo-Yu, et al.
Veröffentlicht: (2025)
von: Chen, Bo-Yu, et al.
Veröffentlicht: (2025)
Neuromorphic Principles for Efficient Large Language Models on Intel Loihi 2
von: Abreu, Steven, et al.
Veröffentlicht: (2025)
von: Abreu, Steven, et al.
Veröffentlicht: (2025)
Evaluation of NVENC Split-Frame Encoding (SFE) for UHD Video Transcoding
von: Arunruangsirilert, Kasidis, et al.
Veröffentlicht: (2025)
von: Arunruangsirilert, Kasidis, et al.
Veröffentlicht: (2025)
Real Time FPGA Based Transformers & VLMs for Vision Tasks: SOTA Designs and Optimizations
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025)
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025)
FireFly-P: FPGA-Accelerated Spiking Neural Network Plasticity for Robust Adaptive Control
von: Li, Tenglong, et al.
Veröffentlicht: (2026)
von: Li, Tenglong, et al.
Veröffentlicht: (2026)
SSRESF: Sensitivity-aware Single-particle Radiation Effects Simulation Framework in SoC Platforms based on SVM Algorithm
von: Liu, Meng, et al.
Veröffentlicht: (2024)
von: Liu, Meng, et al.
Veröffentlicht: (2024)
Real Time FPGA Based CNNs for Detection, Classification, and Tracking in Autonomous Systems: State of the Art Designs and Optimizations
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025)
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025)
FireFly-T: High-Throughput Sparsity Exploitation for Spiking Transformer Acceleration with Dual-Engine Overlay Architecture
von: Li, Tenglong, et al.
Veröffentlicht: (2025)
von: Li, Tenglong, et al.
Veröffentlicht: (2025)
FireFly-S: Exploiting Dual-Side Sparsity for Spiking Neural Networks Acceleration with Reconfigurable Spatial Architecture
von: Li, Tenglong, et al.
Veröffentlicht: (2024)
von: Li, Tenglong, et al.
Veröffentlicht: (2024)
A Fully-Configurable Open-Source Software-Defined Digital Quantized Spiking Neural Core Architecture
von: Matinizadeh, Shadi, et al.
Veröffentlicht: (2024)
von: Matinizadeh, Shadi, et al.
Veröffentlicht: (2024)
Cocco: Hardware-Mapping Co-Exploration towards Memory Capacity-Communication Optimization
von: Tan, Zhanhong, et al.
Veröffentlicht: (2024)
von: Tan, Zhanhong, et al.
Veröffentlicht: (2024)
Evaluation of GPU Video Encoder for Low-Latency Real-Time 4K UHD Encoding
von: Arunruangsirilert, Kasidis, et al.
Veröffentlicht: (2025)
von: Arunruangsirilert, Kasidis, et al.
Veröffentlicht: (2025)
MCFlash: Bulk Bitwise Processing in 3D NAND with Dynamic Sensing and Multi-level Encoding
von: Rahman, Habib Ur, et al.
Veröffentlicht: (2026)
von: Rahman, Habib Ur, et al.
Veröffentlicht: (2026)
When Pipelined In-Memory Accelerators Meet Spiking Direct Feedback Alignment: A Co-Design for Neuromorphic Edge Computing
von: Ren, Haoxiong, et al.
Veröffentlicht: (2025)
von: Ren, Haoxiong, et al.
Veröffentlicht: (2025)
Travel Time Based Task Mapping for NoC-Based DNN Accelerator
von: Chen, Yizhi, et al.
Veröffentlicht: (2024)
von: Chen, Yizhi, et al.
Veröffentlicht: (2024)
Towards Generalized On-Chip Communication for Programmable Accelerators in Heterogeneous Architectures
von: Zuckerman, Joseph, et al.
Veröffentlicht: (2024)
von: Zuckerman, Joseph, et al.
Veröffentlicht: (2024)
UniSpike: Accelerating Spiking Neural Networks on Neuromorphic Systems via Eliminating Address Redundancy
von: Xing, Qinghui, et al.
Veröffentlicht: (2026)
von: Xing, Qinghui, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Flex-TPU: A Flexible TPU with Runtime Reconfigurable Dataflow Architecture
von: Elbtity, Mohammed, et al.
Veröffentlicht: (2024) -
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-bit LLMs
von: Malekar, Jinendra, et al.
Veröffentlicht: (2025) -
TPU-Gen: LLM-Driven Custom Tensor Processing Unit Generator
von: Vungarala, Deepak, et al.
Veröffentlicht: (2025) -
Analytical Heterogeneous Die-to-Die 3D Placement with Macros
von: Zhao, Yuxuan, et al.
Veröffentlicht: (2024) -
CrossNAS: A Cross-Layer Neural Architecture Search Framework for PIM Systems
von: Amin, Md Hasibul, et al.
Veröffentlicht: (2025)