CGRA4ML: A Hardware/Software Framework to Implement Neural Networks for Scientific Edge Computing
Fuente:
arXiv
Saved in:
| Main Authors: | Abarajithan, G, Ma, Zhenghua, Munasinghe, Ravidu, Restuccia, Francesco, Kastner, Ryan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unconventional Universal Computation in Babbage's Analytical Engine
by: Rojas, Raul
Published: (2024)
by: Rojas, Raul
Published: (2024)
FireBridge: Cycle-Accurate Hardware + Firmware Co-Verification for Modern Accelerators
by: Abarajithan, G, et al.
Published: (2026)
by: Abarajithan, G, et al.
Published: (2026)
Computing with Clocks
by: Edwards, Jonathan, et al.
Published: (2024)
by: Edwards, Jonathan, et al.
Published: (2024)
Design and Implementation of a Takum Arithmetic Hardware Codec
by: Hunhold, Laslo
Published: (2024)
by: Hunhold, Laslo
Published: (2024)
Design Rules for Extreme-Edge Scientific Computing on AI Engines
by: Ma, Zhenghua, et al.
Published: (2026)
by: Ma, Zhenghua, et al.
Published: (2026)
SPEC CPU: The Next Generation
by: Madhav, Mahesh, et al.
Published: (2026)
by: Madhav, Mahesh, et al.
Published: (2026)
HG-PIPE: Vision Transformer Acceleration with Hybrid-Grained Pipeline
by: Guo, Qingyu, et al.
Published: (2024)
by: Guo, Qingyu, et al.
Published: (2024)
Roughness and entropy measures of a soft set
by: Acharjee, Santanu, et al.
Published: (2026)
by: Acharjee, Santanu, et al.
Published: (2026)
Recent Advances in Data-Driven Business Process Management
by: Ackermann, Lars, et al.
Published: (2024)
by: Ackermann, Lars, et al.
Published: (2024)
TOM: A Ternary Read-only Memory Accelerator for LLM-powered Edge Intelligence
by: Guan, Hongyi, et al.
Published: (2026)
by: Guan, Hongyi, et al.
Published: (2026)
A Human-In-The-Loop Approach for Improving Fairness in Predictive Business Process Monitoring
by: Käppel, Martin, et al.
Published: (2025)
by: Käppel, Martin, et al.
Published: (2025)
A flexible framework for early power and timing comparison of time-multiplexed CGRA kernel executions
by: Aspros, Maxime Henri, et al.
Published: (2025)
by: Aspros, Maxime Henri, et al.
Published: (2025)
Dataflow & Tiling Strategies in Edge-AI FPGA Accelerators: A Comprehensive Literature Review
by: Li, Richie
Published: (2025)
by: Li, Richie
Published: (2025)
NeuroPDE: A Neuromorphic PDE Solver Based on Spintronic and Ferroelectric Devices
by: Fu, Siqing, et al.
Published: (2025)
by: Fu, Siqing, et al.
Published: (2025)
Design and Implementation of an FPGA-Based Hardware Accelerator for Transformer
by: Li, Richie, et al.
Published: (2025)
by: Li, Richie, et al.
Published: (2025)
Biological Intuition on Digital Hardware: An RTL Implementation of Poisson-Encoded SNNs for Static Image Classification
by: Das, Debabrata, et al.
Published: (2026)
by: Das, Debabrata, et al.
Published: (2026)
DSPE: An Energy-Efficient Edge Processor for DeepSeek Inference with MerkleTree-based Incremental Pruning, Multi-Stage Boothing Lookup and Dynamic Adaptive Posit Processing
by: Zhang, Yuhan, et al.
Published: (2026)
by: Zhang, Yuhan, et al.
Published: (2026)
Attention Please: What Transformer Models Really Learn for Process Prediction
by: Käppel, Martin, et al.
Published: (2024)
by: Käppel, Martin, et al.
Published: (2024)
Socially Beneficial Metaverse: Framework, Technologies, Applications, and Challenges
by: Xu, Xiaolong, et al.
Published: (2023)
by: Xu, Xiaolong, et al.
Published: (2023)
Graph Transformers: A Survey
by: Shehzad, Ahsan, et al.
Published: (2024)
by: Shehzad, Ahsan, et al.
Published: (2024)
Subspace Clustering in Wavelet Packets Domain
by: Kopriva, Ivica, et al.
Published: (2024)
by: Kopriva, Ivica, et al.
Published: (2024)
Bare-Metal Tensor Virtualization: Overcoming the Memory Wall in Edge-AI Inference on ARM64
by: Kilictas, Bugra, et al.
Published: (2026)
by: Kilictas, Bugra, et al.
Published: (2026)
RISC-V Based TinyML Accelerator for Depthwise Separable Convolutions in Edge AI
by: Yildirim, Muhammed, et al.
Published: (2025)
by: Yildirim, Muhammed, et al.
Published: (2025)
basic_RV32s: An Open-Source Microarchitectural Roadmap for RISC-V RV32I
by: Kang, Hyun Woo, et al.
Published: (2025)
by: Kang, Hyun Woo, et al.
Published: (2025)
Accelerating State-Vector Quantum Simulation on Integrated GPUs via Cache Locality Optimization: A Cross-Architecture Evaluation
by: Thomaz, Gabriel Fernandes, et al.
Published: (2026)
by: Thomaz, Gabriel Fernandes, et al.
Published: (2026)
Authenticated Delegation and Authorized AI Agents
by: South, Tobin, et al.
Published: (2025)
by: South, Tobin, et al.
Published: (2025)
Towards LLM-based Generation of Human-Readable Proofs in Polynomial Formal Verification
by: Drechsler, Rolf
Published: (2025)
by: Drechsler, Rolf
Published: (2025)
The Selective G-Bispectrum and its Inversion: Applications to G-Invariant Networks
by: Mataigne, Simon, et al.
Published: (2024)
by: Mataigne, Simon, et al.
Published: (2024)
A survey on FPGA-based accelerator for ML models
by: Yan, Feng, et al.
Published: (2024)
by: Yan, Feng, et al.
Published: (2024)
The $qs$ Inequality: Quantifying the Double Penalty of Mixture-of-Experts at Inference
by: Adhinarayanan, Vignesh, et al.
Published: (2026)
by: Adhinarayanan, Vignesh, et al.
Published: (2026)
Complete Reduction for Derivatives in a Transcendental Liouvillian Extension
by: Chen, Shaoshi, et al.
Published: (2026)
by: Chen, Shaoshi, et al.
Published: (2026)
Complete Reduction for Derivatives in a Primitive Tower
by: Du, Hao, et al.
Published: (2025)
by: Du, Hao, et al.
Published: (2025)
Formalising CXL Cache Coherence
by: Tan, Chengsong, et al.
Published: (2024)
by: Tan, Chengsong, et al.
Published: (2024)
HERCULES: Hardware-Efficient, Robust, Continual Learning Neural Architecture Search
by: Gambella, Matteo, et al.
Published: (2026)
by: Gambella, Matteo, et al.
Published: (2026)
Efficient Fine-Tuning Methods for Portuguese Question Answering: A Comparative Study of PEFT on BERTimbau and Exploratory Evaluation of Generative LLMs
by: Nina, Mariela M., et al.
Published: (2026)
by: Nina, Mariela M., et al.
Published: (2026)
Glass-Box Analysis for Computer Systems: Transparency Index, Shapley Attribution, and Markov Models of Branch Prediction
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
EatGAN: An Edge-Attention Guided Generative Adversarial Network for Single Image Super-Resolution
by: Rao, Penghao, et al.
Published: (2025)
by: Rao, Penghao, et al.
Published: (2025)
RV-IM100: Quantifying ISA Extension, Datapath Width, and Pipeline Depth Trade-offs in RISC-V Microarchitectures
by: Kang, Hyunwoo
Published: (2026)
by: Kang, Hyunwoo
Published: (2026)
Time Domain Near Memory Computing Engine
by: Antal, Sarthak, et al.
Published: (2026)
by: Antal, Sarthak, et al.
Published: (2026)
The Monte Carlo Method and New Device and Architectural Techniques for Accelerating It
by: Petangoda, Janith, et al.
Published: (2025)
by: Petangoda, Janith, et al.
Published: (2025)
Similar Items
-
Unconventional Universal Computation in Babbage's Analytical Engine
by: Rojas, Raul
Published: (2024) -
FireBridge: Cycle-Accurate Hardware + Firmware Co-Verification for Modern Accelerators
by: Abarajithan, G, et al.
Published: (2026) -
Computing with Clocks
by: Edwards, Jonathan, et al.
Published: (2024) -
Design and Implementation of a Takum Arithmetic Hardware Codec
by: Hunhold, Laslo
Published: (2024) -
Design Rules for Extreme-Edge Scientific Computing on AI Engines
by: Ma, Zhenghua, et al.
Published: (2026)