Reuse and Blend: Energy-Efficient Optical Neural Network Enabled by Weight Sharing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Bo, Fang, Yuetong, Yu, Shaoliang, Xu, Renjing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Layer-wise Weight Selection for Power-Efficient Neural Network Acceleration
von: Fang, Jiaxun, et al.
Veröffentlicht: (2025)
von: Fang, Jiaxun, et al.
Veröffentlicht: (2025)
Enabling Efficient Transaction Processing on CXL-Based Memory Sharing
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
von: Wang, Zhao, et al.
Veröffentlicht: (2025)
Transitive Array: An Efficient GEMM Accelerator with Result Reuse
von: Guo, Cong, et al.
Veröffentlicht: (2025)
von: Guo, Cong, et al.
Veröffentlicht: (2025)
NeuroBlend: Towards Low-Power yet Accurate Neural Network-Based Inference Engine Blending Binary and Fixed-Point Convolutions
von: Fayyazi, Arash, et al.
Veröffentlicht: (2023)
von: Fayyazi, Arash, et al.
Veröffentlicht: (2023)
Energy-Efficient FPGA Framework for Non-Quantized Convolutional Neural Networks
von: Athanasiadis, Angelos, et al.
Veröffentlicht: (2025)
von: Athanasiadis, Angelos, et al.
Veröffentlicht: (2025)
A 16 nm 1.60TOPS/W High Utilization DNN Accelerator with 3D Spatial Data Reuse and Efficient Shared Memory Access
von: Yi, Xiaoling, et al.
Veröffentlicht: (2026)
von: Yi, Xiaoling, et al.
Veröffentlicht: (2026)
SemanticBBV: A Semantic Signature for Cross-Program Knowledge Reuse in Microarchitecture Simulation
von: Liu, Zhenguo, et al.
Veröffentlicht: (2025)
von: Liu, Zhenguo, et al.
Veröffentlicht: (2025)
Enabling Efficient Hybrid Systolic Computation in Shared L1-Memory Manycore Clusters
von: Mazzola, Sergio, et al.
Veröffentlicht: (2024)
von: Mazzola, Sergio, et al.
Veröffentlicht: (2024)
Algorithmic Strategies for Sustainable Reuse of Neural Network Accelerators with Permanent Faults
von: Alama, Youssef A. Ait, et al.
Veröffentlicht: (2024)
von: Alama, Youssef A. Ait, et al.
Veröffentlicht: (2024)
ROSA: Robust and Energy-Efficient Microring-Based Optical Neural Networks via Optical Shift-and-Add and Layer-Wise Hybrid Mapping
von: Zhang, Huifan, et al.
Veröffentlicht: (2026)
von: Zhang, Huifan, et al.
Veröffentlicht: (2026)
Sparse-on-Dense: Area and Energy-Efficient Computing of Sparse Neural Networks on Dense Matrix Multiplication Accelerators
von: Yoon, Hyunsung, et al.
Veröffentlicht: (2026)
von: Yoon, Hyunsung, et al.
Veröffentlicht: (2026)
A PVT-Resilient Subthreshold SRAM-Based In-Memory Computing Accelerator with In-Situ Regulation for Energy-Efficient Spiking Neural Networks
von: Kao, Shih-Hang, et al.
Veröffentlicht: (2026)
von: Kao, Shih-Hang, et al.
Veröffentlicht: (2026)
CIMPool: Scalable Neural Network Acceleration for Compute-In-Memory using Weight Pools
von: Li, Shurui, et al.
Veröffentlicht: (2025)
von: Li, Shurui, et al.
Veröffentlicht: (2025)
The Immutable Tensor Architecture: A Pure Dataflow Approach for Secure, Energy-Efficient AI Inference
von: Li, Fang
Veröffentlicht: (2025)
von: Li, Fang
Veröffentlicht: (2025)
TsetlinKWS: A 65nm 16.58uW, 0.63mm2 State-Driven Convolutional Tsetlin Machine-Based Accelerator For Keyword Spotting
von: Lin, Baizhou, et al.
Veröffentlicht: (2025)
von: Lin, Baizhou, et al.
Veröffentlicht: (2025)
Hardware-Aware Neural Network Compilation with Learned Optimization: A RISC-V Accelerator Approach
von: Ganti, Ravindra, et al.
Veröffentlicht: (2025)
von: Ganti, Ravindra, et al.
Veröffentlicht: (2025)
High Utilization Energy-Aware Real-Time Inference Deep Convolutional Neural Network Accelerator
von: Lin, Kuan-Ting, et al.
Veröffentlicht: (2025)
von: Lin, Kuan-Ting, et al.
Veröffentlicht: (2025)
Reuse Detector: Improving the Management of STT-RAM SLLCs
von: RodrÍguez-RodrÍguez, Roberto, et al.
Veröffentlicht: (2024)
von: RodrÍguez-RodrÍguez, Roberto, et al.
Veröffentlicht: (2024)
Enabling Efficient Hardware Acceleration of Hybrid Vision Transformer (ViT) Networks at the Edge
von: Dumoulin, Joren, et al.
Veröffentlicht: (2025)
von: Dumoulin, Joren, et al.
Veröffentlicht: (2025)
Shared-PIM: Enabling Concurrent Computation and Data Flow for Faster Processing-in-DRAM
von: Mamdouh, Ahmed, et al.
Veröffentlicht: (2024)
von: Mamdouh, Ahmed, et al.
Veröffentlicht: (2024)
A Flexible Precision Scaling Deep Neural Network Accelerator with Efficient Weight Combination
von: Zhao, Liang, et al.
Veröffentlicht: (2025)
von: Zhao, Liang, et al.
Veröffentlicht: (2025)
HyDRA: Deadline and Reuse-Aware Cacheability for Hardware Accelerators
von: Agarwal, Ayushi, et al.
Veröffentlicht: (2026)
von: Agarwal, Ayushi, et al.
Veröffentlicht: (2026)
A Scalable RISC-V Vector Processor Enabling Efficient Multi-Precision DNN Inference
von: Wang, Chuanning, et al.
Veröffentlicht: (2024)
von: Wang, Chuanning, et al.
Veröffentlicht: (2024)
SPEED: A Scalable RISC-V Vector Processor Enabling Efficient Multi-Precision DNN Inference
von: Wang, Chuanning, et al.
Veröffentlicht: (2024)
von: Wang, Chuanning, et al.
Veröffentlicht: (2024)
FETTA: Flexible and Efficient Hardware Accelerator for Tensorized Neural Network Training
von: Lu, Jinming, et al.
Veröffentlicht: (2025)
von: Lu, Jinming, et al.
Veröffentlicht: (2025)
Energy-Efficient Hardware Acceleration of Whisper ASR on a CGLA
von: Ando, Takuto, et al.
Veröffentlicht: (2025)
von: Ando, Takuto, et al.
Veröffentlicht: (2025)
Systolic Array Data Flows for Efficient Matrix Multiplication in Deep Neural Networks
von: Raja, Tejas
Veröffentlicht: (2024)
von: Raja, Tejas
Veröffentlicht: (2024)
Efficient Nonlinear Function Approximation in Analog Resistive Crossbars for Recurrent Neural Networks
von: Yang, Junyi, et al.
Veröffentlicht: (2024)
von: Yang, Junyi, et al.
Veröffentlicht: (2024)
A Low-Power Sparse Deep Learning Accelerator with Optimized Data Reuse
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2025)
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2025)
RACAM: Enhancing DRAM with Reuse-Aware Computation and Automated Mapping for ML Inference
von: Ma, Siyuan, et al.
Veröffentlicht: (2025)
von: Ma, Siyuan, et al.
Veröffentlicht: (2025)
A Bit Level Weight Reordering Strategy Based on Column Similarity to Explore Weight Sparsity in RRAM-based NN Accelerator
von: Yang, Weiping, et al.
Veröffentlicht: (2025)
von: Yang, Weiping, et al.
Veröffentlicht: (2025)
A Switch-Centric In-Network Architecture for Accelerating LLM Inference in Shared-Memory Network
von: Jiang, Aojie, et al.
Veröffentlicht: (2026)
von: Jiang, Aojie, et al.
Veröffentlicht: (2026)
Hardware Efficient Accelerator for Spiking Transformer With Reconfigurable Parallel Time Step Computing
von: Chen, Bo-Yu, et al.
Veröffentlicht: (2025)
von: Chen, Bo-Yu, et al.
Veröffentlicht: (2025)
SSD Offloading for LLM Mixture-of-Experts Weights Considered Harmful in Energy Efficiency
von: Kyung, Kwanhee, et al.
Veröffentlicht: (2025)
von: Kyung, Kwanhee, et al.
Veröffentlicht: (2025)
An Efficient Data Reuse with Tile-Based Adaptive Stationary for Transformer Accelerators
von: Li, Tseng-Jen, et al.
Veröffentlicht: (2025)
von: Li, Tseng-Jen, et al.
Veröffentlicht: (2025)
GTA: a new General Tensor Accelerator with Better Area Efficiency and Data Reuse
von: Ai, Chenyang, et al.
Veröffentlicht: (2024)
von: Ai, Chenyang, et al.
Veröffentlicht: (2024)
ONE-SA: Enabling Nonlinear Operations in Systolic Arrays for Efficient and Flexible Neural Network Inference
von: Sun, Ruiqi, et al.
Veröffentlicht: (2024)
von: Sun, Ruiqi, et al.
Veröffentlicht: (2024)
Voxel-CIM: An Efficient Compute-in-Memory Accelerator for Voxel-based Point Cloud Neural Networks
von: Lin, Xipeng, et al.
Veröffentlicht: (2024)
von: Lin, Xipeng, et al.
Veröffentlicht: (2024)
A Logic-Reuse Approach to Nibble-based Multiplier Design for Low Power Vector Computing
von: Chowdhury, Md Rownak Hossain, et al.
Veröffentlicht: (2026)
von: Chowdhury, Md Rownak Hossain, et al.
Veröffentlicht: (2026)
SegSEM: Enabling and Enhancing SAM2 for SEM Contour Extraction
von: Chen, Da, et al.
Veröffentlicht: (2026)
von: Chen, Da, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Layer-wise Weight Selection for Power-Efficient Neural Network Acceleration
von: Fang, Jiaxun, et al.
Veröffentlicht: (2025) -
Enabling Efficient Transaction Processing on CXL-Based Memory Sharing
von: Wang, Zhao, et al.
Veröffentlicht: (2025) -
Transitive Array: An Efficient GEMM Accelerator with Result Reuse
von: Guo, Cong, et al.
Veröffentlicht: (2025) -
NeuroBlend: Towards Low-Power yet Accurate Neural Network-Based Inference Engine Blending Binary and Fixed-Point Convolutions
von: Fayyazi, Arash, et al.
Veröffentlicht: (2023) -
Energy-Efficient FPGA Framework for Non-Quantized Convolutional Neural Networks
von: Athanasiadis, Angelos, et al.
Veröffentlicht: (2025)