Holistic Optimization Framework for FPGA Accelerators
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pouget, Stéphane, Lo, Michael, Pouchet, Louis-Noël, Cong, Jason |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Automatic Hardware Pragma Insertion in High-Level Synthesis: A Non-Linear Programming Approach
von: Pouget, Stéphane, et al.
Veröffentlicht: (2024)
von: Pouget, Stéphane, et al.
Veröffentlicht: (2024)
Demystifying FPGA Hard NoC Performance
von: Liu, Sihao, et al.
Veröffentlicht: (2025)
von: Liu, Sihao, et al.
Veröffentlicht: (2025)
A Reconfigurable Framework for AI-FPGA Agent Integration and Acceleration
von: Yunusoglu, Aybars, et al.
Veröffentlicht: (2026)
von: Yunusoglu, Aybars, et al.
Veröffentlicht: (2026)
Implementation and Analysis of Thermometer Encoding in DWN FPGA Accelerators
von: Mecik, Michael, et al.
Veröffentlicht: (2025)
von: Mecik, Michael, et al.
Veröffentlicht: (2025)
An Optimizing Framework on MLIR for Efficient FPGA-based Accelerator Generation
von: Zhang, Weichuang, et al.
Veröffentlicht: (2024)
von: Zhang, Weichuang, et al.
Veröffentlicht: (2024)
Stream-HLS: Towards Automatic Dataflow Acceleration
von: Basalama, Suhail, et al.
Veröffentlicht: (2025)
von: Basalama, Suhail, et al.
Veröffentlicht: (2025)
An Irredundant and Compressed Data Layout to Optimize Bandwidth Utilization of FPGA Accelerators
von: Ferry, Corentin, et al.
Veröffentlicht: (2024)
von: Ferry, Corentin, et al.
Veröffentlicht: (2024)
LaZagna: An Open-Source Framework for Flexible 3D FPGA Architectural Exploration
von: Youssef, Ismael, et al.
Veröffentlicht: (2025)
von: Youssef, Ismael, et al.
Veröffentlicht: (2025)
Swift: A Multi-FPGA Framework for Scaling Up Accelerated Graph Analytics
von: Jaiyeoba, Oluwole, et al.
Veröffentlicht: (2024)
von: Jaiyeoba, Oluwole, et al.
Veröffentlicht: (2024)
SIRA: Scaled-Integer Range Analysis for Optimizing FPGA Dataflow Neural Network Accelerators
von: Umuroglu, Yaman, et al.
Veröffentlicht: (2025)
von: Umuroglu, Yaman, et al.
Veröffentlicht: (2025)
FPGA-Optimized Hardware Accelerator for Fast Fourier Transform and Singular Value Decomposition in AI
von: Ding, Hong, et al.
Veröffentlicht: (2025)
von: Ding, Hong, et al.
Veröffentlicht: (2025)
Bombyx: OpenCilk Compilation for FPGA Hardware Acceleration
von: Shahawy, Mohamed, et al.
Veröffentlicht: (2025)
von: Shahawy, Mohamed, et al.
Veröffentlicht: (2025)
SpecMamba: Accelerating Mamba Inference on FPGA with Speculative Decoding
von: Zhong, Linfeng, et al.
Veröffentlicht: (2025)
von: Zhong, Linfeng, et al.
Veröffentlicht: (2025)
A Novel FPGA-based CNN Hardware Accelerator: Optimization for Convolutional Layers using Karatsuba Ofman Multiplier
von: Sarkar, Amit
Veröffentlicht: (2024)
von: Sarkar, Amit
Veröffentlicht: (2024)
RealProbe: An Automated and Lightweight Performance Profiler for In-FPGA Execution of High-Level Synthesis Designs
von: Kim, Jiho, et al.
Veröffentlicht: (2025)
von: Kim, Jiho, et al.
Veröffentlicht: (2025)
ZynqParrot: A Scale-Down Approach to Cycle-Accurate, FPGA-Accelerated Co-Emulation
von: Ruelas-Petrisko, Daniel, et al.
Veröffentlicht: (2025)
von: Ruelas-Petrisko, Daniel, et al.
Veröffentlicht: (2025)
TurboFuzz: FPGA Accelerated Hardware Fuzzing for Processor Agile Verification
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
SuperUROP: An FPGA-Based Spatial Accelerator for Sparse Matrix Operations
von: Parthasarathy, Rishab
Veröffentlicht: (2025)
von: Parthasarathy, Rishab
Veröffentlicht: (2025)
A High-Throughput FPGA Accelerator for Lightweight CNNs With Balanced Dataflow
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2024)
Hummingbird: A Smaller and Faster Large Language Model Accelerator on Embedded FPGA
von: Li, Jindong, et al.
Veröffentlicht: (2025)
von: Li, Jindong, et al.
Veröffentlicht: (2025)
Facial Expression Recognition System Using DNN Accelerator with Multi-threading on FPGA
von: Ando, Takuto, et al.
Veröffentlicht: (2025)
von: Ando, Takuto, et al.
Veröffentlicht: (2025)
SpeedLLM: An FPGA Co-design of Large Language Model Inference Accelerator
von: Wang, Peipei, et al.
Veröffentlicht: (2025)
von: Wang, Peipei, et al.
Veröffentlicht: (2025)
always_comm: An FPGA-based Hardware Accelerator for Audio/Video Compression and Transmission
von: Parthasarathy, Rishab, et al.
Veröffentlicht: (2025)
von: Parthasarathy, Rishab, et al.
Veröffentlicht: (2025)
FAST-Prefill: FPGA Accelerated Sparse Attention for Long Context LLM Prefill
von: Jayanth, Rakshith, et al.
Veröffentlicht: (2026)
von: Jayanth, Rakshith, et al.
Veröffentlicht: (2026)
HAAN: A Holistic Approach for Accelerating Normalization Operations in Large Language Models
von: Peng, Tianfan, et al.
Veröffentlicht: (2025)
von: Peng, Tianfan, et al.
Veröffentlicht: (2025)
Late Breaking Result: FPGA-Based Emulation and Fault Injection for CNN Inference Accelerators
von: Masar, Filip, et al.
Veröffentlicht: (2025)
von: Masar, Filip, et al.
Veröffentlicht: (2025)
Systolic Sparse Tensor Slices: FPGA Building Blocks for Sparse and Dense AI Acceleration
von: Taka, Endri, et al.
Veröffentlicht: (2025)
von: Taka, Endri, et al.
Veröffentlicht: (2025)
Towards Employing FPGA and ASIP Acceleration to Enable Onboard AI/ML in Space Applications
von: Leon, Vasileios, et al.
Veröffentlicht: (2025)
von: Leon, Vasileios, et al.
Veröffentlicht: (2025)
Graphitron: A Domain Specific Language for FPGA-based Graph Processing Accelerator Generation
von: Zhang, Xinmiao, et al.
Veröffentlicht: (2024)
von: Zhang, Xinmiao, et al.
Veröffentlicht: (2024)
ISAAC: Intelligent, Scalable, Agile, and Accelerated CPU Verification via LLM-aided FPGA Parallelism
von: Sun, Jialin, et al.
Veröffentlicht: (2025)
von: Sun, Jialin, et al.
Veröffentlicht: (2025)
FireFly-P: FPGA-Accelerated Spiking Neural Network Plasticity for Robust Adaptive Control
von: Li, Tenglong, et al.
Veröffentlicht: (2026)
von: Li, Tenglong, et al.
Veröffentlicht: (2026)
Energy-Efficient FPGA Framework for Non-Quantized Convolutional Neural Networks
von: Athanasiadis, Angelos, et al.
Veröffentlicht: (2025)
von: Athanasiadis, Angelos, et al.
Veröffentlicht: (2025)
EdgeLLM: A Highly Efficient CPU-FPGA Heterogeneous Edge Accelerator for Large Language Models
von: Huang, Mingqiang, et al.
Veröffentlicht: (2024)
von: Huang, Mingqiang, et al.
Veröffentlicht: (2024)
Revealing Untapped DSP Optimization Potentials for FPGA-Based Systolic Matrix Engines
von: Li, Jindong, et al.
Veröffentlicht: (2024)
von: Li, Jindong, et al.
Veröffentlicht: (2024)
FERIVer: An FPGA-assisted Emulated Framework for RTL Verification of RISC-V Processors
von: Qin, Kun, et al.
Veröffentlicht: (2025)
von: Qin, Kun, et al.
Veröffentlicht: (2025)
UbiMoE: A Ubiquitous Mixture-of-Experts Vision Transformer Accelerator With Hybrid Computation Pattern on FPGA
von: Dong, Jiale, et al.
Veröffentlicht: (2025)
von: Dong, Jiale, et al.
Veröffentlicht: (2025)
Modeling and Optimizing Performance Bottlenecks for Neuromorphic Accelerators
von: Yik, Jason, et al.
Veröffentlicht: (2025)
von: Yik, Jason, et al.
Veröffentlicht: (2025)
FPGA-Accelerated Correspondence-free Point Cloud Registration with PointNet Features
von: Sugiura, Keisuke, et al.
Veröffentlicht: (2024)
von: Sugiura, Keisuke, et al.
Veröffentlicht: (2024)
Real Time FPGA Based Transformers & VLMs for Vision Tasks: SOTA Designs and Optimizations
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025)
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025)
RePart: Efficient Hypergraph Partitioning with Logic Replication Optimization for Multi-FPGA System
von: Fu, Zizhuo, et al.
Veröffentlicht: (2026)
von: Fu, Zizhuo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Automatic Hardware Pragma Insertion in High-Level Synthesis: A Non-Linear Programming Approach
von: Pouget, Stéphane, et al.
Veröffentlicht: (2024) -
Demystifying FPGA Hard NoC Performance
von: Liu, Sihao, et al.
Veröffentlicht: (2025) -
A Reconfigurable Framework for AI-FPGA Agent Integration and Acceleration
von: Yunusoglu, Aybars, et al.
Veröffentlicht: (2026) -
Implementation and Analysis of Thermometer Encoding in DWN FPGA Accelerators
von: Mecik, Michael, et al.
Veröffentlicht: (2025) -
An Optimizing Framework on MLIR for Efficient FPGA-based Accelerator Generation
von: Zhang, Weichuang, et al.
Veröffentlicht: (2024)