OpenEye: A Scalable Open-Source Hardware Accelerator for DNNs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lebold, Denis, Wöhrle, Hendrik |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FPGA-Accelerated RISC-V ISA Extensions for Efficient Neural Network Inference on Edge Devices
von: Parameshwara, Arya, et al.
Veröffentlicht: (2025)
von: Parameshwara, Arya, et al.
Veröffentlicht: (2025)
Biological Intuition on Digital Hardware: An RTL Implementation of Poisson-Encoded SNNs for Static Image Classification
von: Das, Debabrata, et al.
Veröffentlicht: (2026)
von: Das, Debabrata, et al.
Veröffentlicht: (2026)
Nonvolatile Charge-Domain Attention with HZO Ferroelectric Capacitors: A Simulation-Based Device-to-System Evaluation
von: Abouagour, Faris
Veröffentlicht: (2026)
von: Abouagour, Faris
Veröffentlicht: (2026)
SynapticCore-X: A Modular Neural Processing Architecture for Low-Cost FPGA Acceleration
von: Parameshwara, Arya
Veröffentlicht: (2025)
von: Parameshwara, Arya
Veröffentlicht: (2025)
FlexiBit: Fully Flexible Precision Bit-parallel Accelerator Architecture for Arbitrary Mixed Precision AI
von: Tahmasebi, Faraz, et al.
Veröffentlicht: (2024)
von: Tahmasebi, Faraz, et al.
Veröffentlicht: (2024)
Differentiable Logic Synthesis: Spectral Coefficient Selection via Sinkhorn-Constrained Composition
von: Pavlov, Gorgi
Veröffentlicht: (2026)
von: Pavlov, Gorgi
Veröffentlicht: (2026)
GainSight: A Unified Framework for Data Lifetime Profiling and Heterogeneous Memory Composition
von: Li, Peijing, et al.
Veröffentlicht: (2025)
von: Li, Peijing, et al.
Veröffentlicht: (2025)
CORE: Constraint-Aware One-Step Reinforcement Learning for Simulation-Guided Neural Network Accelerator Design
von: Xiao, Yifeng, et al.
Veröffentlicht: (2025)
von: Xiao, Yifeng, et al.
Veröffentlicht: (2025)
Coflex: Enhancing HW-NAS with Sparse Gaussian Processes for Efficient and Scalable DNN Accelerator Design
von: Ma, Yinhui, et al.
Veröffentlicht: (2025)
von: Ma, Yinhui, et al.
Veröffentlicht: (2025)
GraphPerf-RT: A Graph-Driven Performance Model for Hardware-Aware Scheduling of OpenMP Codes
von: Pivezhandi, Mohammad, et al.
Veröffentlicht: (2025)
von: Pivezhandi, Mohammad, et al.
Veröffentlicht: (2025)
Multi-diseases detection with memristive system on chip
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
SISA: A Scale-In Systolic Array for GEMM Acceleration
von: Altamura, Luigi, et al.
Veröffentlicht: (2026)
von: Altamura, Luigi, et al.
Veröffentlicht: (2026)
MEDEA: A Design-Time Multi-Objective Manager for Energy-Efficient DNN Inference on Heterogeneous Ultra-Low Power Platforms
von: Taji, Hossein, et al.
Veröffentlicht: (2025)
von: Taji, Hossein, et al.
Veröffentlicht: (2025)
FREESS: An Educational Simulator of a RISC-V-Inspired Superscalar Processor Based on Tomasulo's Algorithm
von: Giorgi, Roberto
Veröffentlicht: (2025)
von: Giorgi, Roberto
Veröffentlicht: (2025)
RayFlex: An Open-Source RTL Implementation of the Hardware Ray Tracer Datapath
von: Shen, Fangjia, et al.
Veröffentlicht: (2024)
von: Shen, Fangjia, et al.
Veröffentlicht: (2024)
Design and Implementation of a RISC-V SoC with Custom DSP Accelerators for Edge Computing
von: Yadav, Priyanshu
Veröffentlicht: (2025)
von: Yadav, Priyanshu
Veröffentlicht: (2025)
Fast and Practical Strassen's Matrix Multiplication using FPGAs
von: Ahmad, Afzal, et al.
Veröffentlicht: (2024)
von: Ahmad, Afzal, et al.
Veröffentlicht: (2024)
Make LLM Inference Affordable to Everyone: Augmenting GPU Memory with NDP-DIMM
von: Liu, Lian, et al.
Veröffentlicht: (2025)
von: Liu, Lian, et al.
Veröffentlicht: (2025)
HERMES: High-Performance RISC-V Memory Hierarchy for ML Workloads
von: Suryadevara, Pranav
Veröffentlicht: (2025)
von: Suryadevara, Pranav
Veröffentlicht: (2025)
The Monte Carlo Method and New Device and Architectural Techniques for Accelerating It
von: Petangoda, Janith, et al.
Veröffentlicht: (2025)
von: Petangoda, Janith, et al.
Veröffentlicht: (2025)
LLM-Driven Large-Scale Spectrum Access
von: Yang, Ning, et al.
Veröffentlicht: (2026)
von: Yang, Ning, et al.
Veröffentlicht: (2026)
A WASM-Subset Stack Architecture for Low-cost FPGAs using Open-Source EDA Flows
von: Chakrabarti, Aradhya
Veröffentlicht: (2025)
von: Chakrabarti, Aradhya
Veröffentlicht: (2025)
ASTER: Attention-based Spiking Transformer Engine for Event-driven Reasoning
von: Das, Tamoghno, et al.
Veröffentlicht: (2025)
von: Das, Tamoghno, et al.
Veröffentlicht: (2025)
A flexible framework for early power and timing comparison of time-multiplexed CGRA kernel executions
von: Aspros, Maxime Henri, et al.
Veröffentlicht: (2025)
von: Aspros, Maxime Henri, et al.
Veröffentlicht: (2025)
Chameleon: A MatMul-Free Temporal Convolutional Network Accelerator for End-to-End Few-Shot and Continual Learning from Sequential Data
von: Blanken, Douwe den, et al.
Veröffentlicht: (2025)
von: Blanken, Douwe den, et al.
Veröffentlicht: (2025)
Architecting Long-Context LLM Acceleration with Packing-Prefetch Scheduler and Ultra-Large Capacity On-Chip Memories
von: Lee, Ming-Yen, et al.
Veröffentlicht: (2025)
von: Lee, Ming-Yen, et al.
Veröffentlicht: (2025)
Research on LLM Acceleration Using the High-Performance RISC-V Processor "Xiangshan" (Nanhu Version) Based on the Open-Source Matrix Instruction Set Extension (Vector Dot Product)
von: Chen, Xu-Hao, et al.
Veröffentlicht: (2024)
von: Chen, Xu-Hao, et al.
Veröffentlicht: (2024)
Mestra: Exploring Migration on Virtualized CGRAs
von: Kyriazis, Agamemnon, et al.
Veröffentlicht: (2026)
von: Kyriazis, Agamemnon, et al.
Veröffentlicht: (2026)
BMR and BWR: Two simple metaphor-free optimization algorithms for solving real-life non-convex constrained and unconstrained problems
von: Rao, Ravipudi Venkata, et al.
Veröffentlicht: (2024)
von: Rao, Ravipudi Venkata, et al.
Veröffentlicht: (2024)
SambaNova SN40L: Scaling the AI Memory Wall with Dataflow and Composition of Experts
von: Prabhakar, Raghu, et al.
Veröffentlicht: (2024)
von: Prabhakar, Raghu, et al.
Veröffentlicht: (2024)
PoCL-R: An Open Standard Based Offloading Layer for Heterogeneous Multi-Access Edge Computing with Server Side Scalability
von: Solanti, Jan, et al.
Veröffentlicht: (2023)
von: Solanti, Jan, et al.
Veröffentlicht: (2023)
RISC-V Based TinyML Accelerator for Depthwise Separable Convolutions in Edge AI
von: Yildirim, Muhammed, et al.
Veröffentlicht: (2025)
von: Yildirim, Muhammed, et al.
Veröffentlicht: (2025)
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation
von: Taha, Hazem, et al.
Veröffentlicht: (2025)
von: Taha, Hazem, et al.
Veröffentlicht: (2025)
Plug-and-Play Spiking Operators: Breaking the Nonlinearity Bottleneck in Spiking Transformers
von: Yuan, Xinzhe, et al.
Veröffentlicht: (2026)
von: Yuan, Xinzhe, et al.
Veröffentlicht: (2026)
DSPE: An Energy-Efficient Edge Processor for DeepSeek Inference with MerkleTree-based Incremental Pruning, Multi-Stage Boothing Lookup and Dynamic Adaptive Posit Processing
von: Zhang, Yuhan, et al.
Veröffentlicht: (2026)
von: Zhang, Yuhan, et al.
Veröffentlicht: (2026)
What Scalable Second-Order Information Knows for Pruning at Initialization
von: Navarrete, Ivo Gollini, et al.
Veröffentlicht: (2025)
von: Navarrete, Ivo Gollini, et al.
Veröffentlicht: (2025)
ChipCraftBrain: Validation-First RTL Generation via Multi-Agent Orchestration
von: Eryilmaz, Cagri
Veröffentlicht: (2026)
von: Eryilmaz, Cagri
Veröffentlicht: (2026)
TokenStack: A Heterogeneous HBM-PIM Architecture and Runtime for Efficient LLM Inference
von: Li, Zhuoran, et al.
Veröffentlicht: (2026)
von: Li, Zhuoran, et al.
Veröffentlicht: (2026)
Ember: A Compiler for Efficient Embedding Operations on Decoupled Access-Execute Architectures
von: Siracusa, Marco, et al.
Veröffentlicht: (2025)
von: Siracusa, Marco, et al.
Veröffentlicht: (2025)
Machine Learning for Energy-Performance-aware Scheduling
von: Hu, Zheyuan, et al.
Veröffentlicht: (2026)
von: Hu, Zheyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
FPGA-Accelerated RISC-V ISA Extensions for Efficient Neural Network Inference on Edge Devices
von: Parameshwara, Arya, et al.
Veröffentlicht: (2025) -
Biological Intuition on Digital Hardware: An RTL Implementation of Poisson-Encoded SNNs for Static Image Classification
von: Das, Debabrata, et al.
Veröffentlicht: (2026) -
Nonvolatile Charge-Domain Attention with HZO Ferroelectric Capacitors: A Simulation-Based Device-to-System Evaluation
von: Abouagour, Faris
Veröffentlicht: (2026) -
SynapticCore-X: A Modular Neural Processing Architecture for Low-Cost FPGA Acceleration
von: Parameshwara, Arya
Veröffentlicht: (2025) -
FlexiBit: Fully Flexible Precision Bit-parallel Accelerator Architecture for Arbitrary Mixed Precision AI
von: Tahmasebi, Faraz, et al.
Veröffentlicht: (2024)