Sustainable Transformer Neural Network Acceleration with Stochastic Photonic Computing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Afifi, S., Alo, O., Thakkar, I., Pasricha, S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ARTEMIS: A Mixed Analog-Stochastic In-DRAM Accelerator for Transformer Neural Networks
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
SafeLight: Enhancing Security in Optical Convolutional Neural Network Accelerators
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
PhotoGAN: Generative Adversarial Neural Network Acceleration with Silicon Photonics
von: Suresh, Tharini, et al.
Veröffentlicht: (2025)
von: Suresh, Tharini, et al.
Veröffentlicht: (2025)
Accelerating Neural Networks for Large Language Models and Graph Processing with Silicon Photonics
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
von: Afifi, Salma, et al.
Veröffentlicht: (2024)
Accelerating Diffusion Models for Generative AI Applications with Silicon Photonics
von: Suresh, Tharini, et al.
Veröffentlicht: (2026)
von: Suresh, Tharini, et al.
Veröffentlicht: (2026)
Optical Computing for Deep Neural Network Acceleration: Foundations, Recent Developments, and Emerging Directions
von: Pasricha, Sudeep
Veröffentlicht: (2024)
von: Pasricha, Sudeep
Veröffentlicht: (2024)
Scaling Photonic Tensor Cores with Unary and Homodyne Designs
von: Alo, Oluwaseun, et al.
Veröffentlicht: (2026)
von: Alo, Oluwaseun, et al.
Veröffentlicht: (2026)
A Comparative Analysis of Microrings Based Incoherent Photonic GEMM Accelerators
von: Vatsavai, Sairam Sri, et al.
Veröffentlicht: (2024)
von: Vatsavai, Sairam Sri, et al.
Veröffentlicht: (2024)
Silicon Photonic 2.5D Interposer Networks for Overcoming Communication Bottlenecks in Scale-out Machine Learning Hardware Accelerators
von: Sunny, Febin, et al.
Veröffentlicht: (2024)
von: Sunny, Febin, et al.
Veröffentlicht: (2024)
OPIMA: Optical Processing-In-Memory for Convolutional Neural Network Acceleration
von: Sunny, Febin, et al.
Veröffentlicht: (2024)
von: Sunny, Febin, et al.
Veröffentlicht: (2024)
Scaling Analog Photonic Accelerators for Byte-Size, Integer General Matrix Multiply (GEMM) Kernels
von: Alo, Oluwaseun Adewunmi, et al.
Veröffentlicht: (2024)
von: Alo, Oluwaseun Adewunmi, et al.
Veröffentlicht: (2024)
Algorithmic Strategies for Sustainable Reuse of Neural Network Accelerators with Permanent Faults
von: Alama, Youssef A. Ait, et al.
Veröffentlicht: (2024)
von: Alama, Youssef A. Ait, et al.
Veröffentlicht: (2024)
An Analog and Digital Hybrid Attention Accelerator for Transformers with Charge-based In-memory Computing
von: Moradifirouzabadi, Ashkan, et al.
Veröffentlicht: (2024)
von: Moradifirouzabadi, Ashkan, et al.
Veröffentlicht: (2024)
U-SWIM: Universal Selective Write-Verify for Computing-in-Memory Neural Accelerators
von: Yan, Zheyu, et al.
Veröffentlicht: (2023)
von: Yan, Zheyu, et al.
Veröffentlicht: (2023)
Accelerating Sparse Graph Neural Networks with Tensor Core Optimization
von: Wu, Ka Wai
Veröffentlicht: (2024)
von: Wu, Ka Wai
Veröffentlicht: (2024)
ESACT: An End-to-End Sparse Accelerator for Compute-Intensive Transformers via Local Similarity
von: Liu, Hongxiang, et al.
Veröffentlicht: (2025)
von: Liu, Hongxiang, et al.
Veröffentlicht: (2025)
Hardware-Efficient Photonic Tensor Core: Accelerating Deep Neural Networks with Structured Compression
von: Ning, Shupeng, et al.
Veröffentlicht: (2025)
von: Ning, Shupeng, et al.
Veröffentlicht: (2025)
A Runtime-Adaptive Transformer Neural Network Accelerator on FPGAs
von: Kabir, Ehsan, et al.
Veröffentlicht: (2024)
von: Kabir, Ehsan, et al.
Veröffentlicht: (2024)
VEXP: A Low-Cost RISC-V ISA Extension for Accelerated Softmax Computation in Transformers
von: Wang, Run, et al.
Veröffentlicht: (2025)
von: Wang, Run, et al.
Veröffentlicht: (2025)
Layer-wise Weight Selection for Power-Efficient Neural Network Acceleration
von: Fang, Jiaxun, et al.
Veröffentlicht: (2025)
von: Fang, Jiaxun, et al.
Veröffentlicht: (2025)
CAFEEN: A Cooperative Approach for Energy Efficient NoCs with Multi-Agent Reinforcement Learning
von: Khan, Kamil, et al.
Veröffentlicht: (2024)
von: Khan, Kamil, et al.
Veröffentlicht: (2024)
When Small Variations Become Big Failures: Reliability Challenges in Compute-in-Memory Neural Accelerators
von: Qin, Yifan, et al.
Veröffentlicht: (2026)
von: Qin, Yifan, et al.
Veröffentlicht: (2026)
Exploring Quantization and Mapping Synergy in Hardware-Aware Deep Neural Network Accelerators
von: Klhufek, Jan, et al.
Veröffentlicht: (2024)
von: Klhufek, Jan, et al.
Veröffentlicht: (2024)
Enhancing Reliability of Neural Networks at the Edge: Inverted Normalization with Stochastic Affine Transformations
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2024)
von: Ahmed, Soyed Tuhin, et al.
Veröffentlicht: (2024)
The prediction of the quality of results in Logic Synthesis using Transformer and Graph Neural Networks
von: Yang, Chenghao, et al.
Veröffentlicht: (2022)
von: Yang, Chenghao, et al.
Veröffentlicht: (2022)
Accelerating Computer Architecture Simulation through Machine Learning
von: Ali, Wajid, et al.
Veröffentlicht: (2024)
von: Ali, Wajid, et al.
Veröffentlicht: (2024)
TreeLUT: An Efficient Alternative to Deep Neural Networks for Inference Acceleration Using Gradient Boosted Decision Trees
von: Khataei, Alireza, et al.
Veröffentlicht: (2025)
von: Khataei, Alireza, et al.
Veröffentlicht: (2025)
End-to-End Transformer Acceleration Through Processing-in-Memory Architectures
von: Yang, Xiaoxuan, et al.
Veröffentlicht: (2025)
von: Yang, Xiaoxuan, et al.
Veröffentlicht: (2025)
ITA: An Energy-Efficient Attention and Softmax Accelerator for Quantized Transformers
von: İslamoğlu, Gamze, et al.
Veröffentlicht: (2023)
von: İslamoğlu, Gamze, et al.
Veröffentlicht: (2023)
Torch2Chip: An End-to-end Customizable Deep Neural Network Compression and Deployment Toolkit for Prototype Hardware Accelerator Design
von: Meng, Jian, et al.
Veröffentlicht: (2024)
von: Meng, Jian, et al.
Veröffentlicht: (2024)
An Efficient Data Reuse with Tile-Based Adaptive Stationary for Transformer Accelerators
von: Li, Tseng-Jen, et al.
Veröffentlicht: (2025)
von: Li, Tseng-Jen, et al.
Veröffentlicht: (2025)
GFormer: Accelerating Large Language Models with Optimized Transformers on Gaudi Processors
von: Zhang, Chengming, et al.
Veröffentlicht: (2024)
von: Zhang, Chengming, et al.
Veröffentlicht: (2024)
An FPGA-Based Reconfigurable Accelerator for Convolution-Transformer Hybrid EfficientViT
von: Shao, Haikuo, et al.
Veröffentlicht: (2024)
von: Shao, Haikuo, et al.
Veröffentlicht: (2024)
iEEG Seizure Detection with a Sparse Hyperdimensional Computing Accelerator
von: Cuyckens, Stef, et al.
Veröffentlicht: (2025)
von: Cuyckens, Stef, et al.
Veröffentlicht: (2025)
Reconfigurable Computing Challenge: Real-Time Graph Neural Networks for Online Event Selection in Big Science
von: Neu, Marc, et al.
Veröffentlicht: (2026)
von: Neu, Marc, et al.
Veröffentlicht: (2026)
TransAxx: Efficient Transformers with Approximate Computing
von: Danopoulos, Dimitrios, et al.
Veröffentlicht: (2024)
von: Danopoulos, Dimitrios, et al.
Veröffentlicht: (2024)
GCoD: Graph Convolutional Network Acceleration via Dedicated Algorithm and Accelerator Co-Design
von: You, Haoran, et al.
Veröffentlicht: (2021)
von: You, Haoran, et al.
Veröffentlicht: (2021)
PiC-BNN: A 128-kbit 65 nm Processing-in-CAM-Based End-to-End Binary Neural Network Accelerator
von: Harary, Yuval, et al.
Veröffentlicht: (2026)
von: Harary, Yuval, et al.
Veröffentlicht: (2026)
COBRA: Algorithm-Architecture Co-optimized Binary Transformer Accelerator for Edge Inference
von: Qiao, Ye, et al.
Veröffentlicht: (2025)
von: Qiao, Ye, et al.
Veröffentlicht: (2025)
Low Power Vision Transformer Accelerator with Hardware-Aware Pruning and Optimized Dataflow
von: Hsiung, Ching-Lin, et al.
Veröffentlicht: (2025)
von: Hsiung, Ching-Lin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ARTEMIS: A Mixed Analog-Stochastic In-DRAM Accelerator for Transformer Neural Networks
von: Afifi, Salma, et al.
Veröffentlicht: (2024) -
SafeLight: Enhancing Security in Optical Convolutional Neural Network Accelerators
von: Afifi, Salma, et al.
Veröffentlicht: (2024) -
PhotoGAN: Generative Adversarial Neural Network Acceleration with Silicon Photonics
von: Suresh, Tharini, et al.
Veröffentlicht: (2025) -
Accelerating Neural Networks for Large Language Models and Graph Processing with Silicon Photonics
von: Afifi, Salma, et al.
Veröffentlicht: (2024) -
Accelerating Diffusion Models for Generative AI Applications with Silicon Photonics
von: Suresh, Tharini, et al.
Veröffentlicht: (2026)