Real Time FPGA Based Transformers & VLMs for Vision Tasks: SOTA Designs and Optimizations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sali, Safa Mohammed, Meribout, Mahmoud, Majeed, Ashiyana Abdul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Real Time FPGA Based CNNs for Detection, Classification, and Tracking in Autonomous Systems: State of the Art Designs and Optimizations
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025)
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025)
Edge GPU Aware Multiple AI Model Pipeline for Accelerated MRI Reconstruction and Analysis
von: Majeed, Ashiyana Abdul, et al.
Veröffentlicht: (2025)
von: Majeed, Ashiyana Abdul, et al.
Veröffentlicht: (2025)
Real-time Object Detection and Associated Hardware Accelerators Targeting Autonomous Vehicles: A Review
von: Sali, Safa, et al.
Veröffentlicht: (2025)
von: Sali, Safa, et al.
Veröffentlicht: (2025)
Scheduling Techniques of AI Models on Modern Heterogeneous Edge GPU -- A Critical Review
von: Majeed, Ashiyana Abdul, et al.
Veröffentlicht: (2025)
von: Majeed, Ashiyana Abdul, et al.
Veröffentlicht: (2025)
Hardware Acceleration in Portable MRIs: State of the Art and Future Prospects
von: Habsi, Omar Al, et al.
Veröffentlicht: (2025)
von: Habsi, Omar Al, et al.
Veröffentlicht: (2025)
Hardware Accelerators for Autonomous Cars: A Review
von: Islayem, Ruba, et al.
Veröffentlicht: (2024)
von: Islayem, Ruba, et al.
Veröffentlicht: (2024)
CoQMoE: Co-Designed Quantization and Computation Orchestration for Mixture-of-Experts Vision Transformer on FPGA
von: Dong, Jiale, et al.
Veröffentlicht: (2025)
von: Dong, Jiale, et al.
Veröffentlicht: (2025)
@NTT: Algorithm-Targeted NTT hardware acceleration via Design-Time Constant Optimization
von: Nabeel, Mohammed, et al.
Veröffentlicht: (2026)
von: Nabeel, Mohammed, et al.
Veröffentlicht: (2026)
Leveraging Simultaneous Usage of Edge GPU Hardware Engines for Video Face Detection and Recognition
von: Baobaid, Asma, et al.
Veröffentlicht: (2025)
von: Baobaid, Asma, et al.
Veröffentlicht: (2025)
Edge-GPU Based Face Tracking for Face Detection and Recognition Acceleration
von: Baobaid, Asma, et al.
Veröffentlicht: (2025)
von: Baobaid, Asma, et al.
Veröffentlicht: (2025)
Blink: Fast Automated Design of Run-Time Power Monitors on FPGA-Based Computing Platforms
von: Galimberti, Andrea, et al.
Veröffentlicht: (2024)
von: Galimberti, Andrea, et al.
Veröffentlicht: (2024)
Design and Implementation of BNN-Based Object Detection on FPGA
von: Zhao, Xuyu, et al.
Veröffentlicht: (2026)
von: Zhao, Xuyu, et al.
Veröffentlicht: (2026)
Real-Time, Energy-Efficient, Sampling-Based Optimal Control via FPGA Acceleration
von: Desai, Tanmay, et al.
Veröffentlicht: (2026)
von: Desai, Tanmay, et al.
Veröffentlicht: (2026)
FPGA-Optimized Hardware Accelerator for Fast Fourier Transform and Singular Value Decomposition in AI
von: Ding, Hong, et al.
Veröffentlicht: (2025)
von: Ding, Hong, et al.
Veröffentlicht: (2025)
RealProbe: An Automated and Lightweight Performance Profiler for In-FPGA Execution of High-Level Synthesis Designs
von: Kim, Jiho, et al.
Veröffentlicht: (2025)
von: Kim, Jiho, et al.
Veröffentlicht: (2025)
Real-Time Adaptive Neural Network on FPGA: Enhancing Adaptability through Dynamic Classifier Selection
von: Bouazzaoui, Achraf El, et al.
Veröffentlicht: (2023)
von: Bouazzaoui, Achraf El, et al.
Veröffentlicht: (2023)
High-Resolution, Multi-Channel FPGA-Based Time-to-Digital Converter
von: Jakli, Balazs, et al.
Veröffentlicht: (2024)
von: Jakli, Balazs, et al.
Veröffentlicht: (2024)
Holistic Optimization Framework for FPGA Accelerators
von: Pouget, Stéphane, et al.
Veröffentlicht: (2025)
von: Pouget, Stéphane, et al.
Veröffentlicht: (2025)
UbiMoE: A Ubiquitous Mixture-of-Experts Vision Transformer Accelerator With Hybrid Computation Pattern on FPGA
von: Dong, Jiale, et al.
Veröffentlicht: (2025)
von: Dong, Jiale, et al.
Veröffentlicht: (2025)
Revealing Untapped DSP Optimization Potentials for FPGA-Based Systolic Matrix Engines
von: Li, Jindong, et al.
Veröffentlicht: (2024)
von: Li, Jindong, et al.
Veröffentlicht: (2024)
FPGA Acceleration of Image Reconstruction for Real-Time Photoacoustic Tomography
von: Gao, Zijian, et al.
Veröffentlicht: (2022)
von: Gao, Zijian, et al.
Veröffentlicht: (2022)
FPGA Resource-aware Structured Pruning for Real-Time Neural Networks
von: Ramhorst, Benjamin, et al.
Veröffentlicht: (2023)
von: Ramhorst, Benjamin, et al.
Veröffentlicht: (2023)
FPPS: An FPGA-Based Point Cloud Processing System
von: Zhou, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Zhou, Xiaofeng, et al.
Veröffentlicht: (2026)
An Irredundant and Compressed Data Layout to Optimize Bandwidth Utilization of FPGA Accelerators
von: Ferry, Corentin, et al.
Veröffentlicht: (2024)
von: Ferry, Corentin, et al.
Veröffentlicht: (2024)
History-Aware Trajectory k-Anonymization Using an FPGA-Based Hardware Accelerator for Real-Time Location Services
von: Nakano, Hiroshi, et al.
Veröffentlicht: (2025)
von: Nakano, Hiroshi, et al.
Veröffentlicht: (2025)
A Composable Dynamic Sparse Dataflow Architecture for Efficient Event-based Vision Processing on FPGA
von: Gao, Yizhao, et al.
Veröffentlicht: (2024)
von: Gao, Yizhao, et al.
Veröffentlicht: (2024)
SuperUROP: An FPGA-Based Spatial Accelerator for Sparse Matrix Operations
von: Parthasarathy, Rishab
Veröffentlicht: (2025)
von: Parthasarathy, Rishab
Veröffentlicht: (2025)
SIRA: Scaled-Integer Range Analysis for Optimizing FPGA Dataflow Neural Network Accelerators
von: Umuroglu, Yaman, et al.
Veröffentlicht: (2025)
von: Umuroglu, Yaman, et al.
Veröffentlicht: (2025)
RePart: Efficient Hypergraph Partitioning with Logic Replication Optimization for Multi-FPGA System
von: Fu, Zizhuo, et al.
Veröffentlicht: (2026)
von: Fu, Zizhuo, et al.
Veröffentlicht: (2026)
Déjà Vu Packing: Optimizing FPGA Logic Clustering Runtime via Pattern Memoization
von: Liebster, Milo, et al.
Veröffentlicht: (2026)
von: Liebster, Milo, et al.
Veröffentlicht: (2026)
An FPGA-Based Reconfigurable Accelerator for Convolution-Transformer Hybrid EfficientViT
von: Shao, Haikuo, et al.
Veröffentlicht: (2024)
von: Shao, Haikuo, et al.
Veröffentlicht: (2024)
Makinote: An FPGA-Based HW/SW Platform for Pre-Silicon Emulation of RISC-V Designs
von: Perdomo, Elias, et al.
Veröffentlicht: (2024)
von: Perdomo, Elias, et al.
Veröffentlicht: (2024)
Practical Timing Closure in FPGA and ASIC Designs: Methods, Challenges, and Case Studies
von: Darvishi, Mostafa
Veröffentlicht: (2025)
von: Darvishi, Mostafa
Veröffentlicht: (2025)
A Scalable FPGA Architecture With Adaptive Memory Utilization for GEMM-Based Operations
von: Petropoulos, Anastasios, et al.
Veröffentlicht: (2025)
von: Petropoulos, Anastasios, et al.
Veröffentlicht: (2025)
Belenos: Bottleneck Evaluation to Link Biomechanics to Novel Computing Optimizations
von: Chitsaz, Hana, et al.
Veröffentlicht: (2025)
von: Chitsaz, Hana, et al.
Veröffentlicht: (2025)
Low-Latency FPGA Control System for Real-Time Neural Network Processing in CCD-Based Trapped-Ion Qubit Measurement
von: Lou, Binglei, et al.
Veröffentlicht: (2025)
von: Lou, Binglei, et al.
Veröffentlicht: (2025)
Travel Time Based Task Mapping for NoC-Based DNN Accelerator
von: Chen, Yizhi, et al.
Veröffentlicht: (2024)
von: Chen, Yizhi, et al.
Veröffentlicht: (2024)
Late Breaking Result: FPGA-Based Emulation and Fault Injection for CNN Inference Accelerators
von: Masar, Filip, et al.
Veröffentlicht: (2025)
von: Masar, Filip, et al.
Veröffentlicht: (2025)
FPGA-Based Multiplier with a New Approximate Full Adder for Error-Resilient Applications
von: Ranjbar, Ali, et al.
Veröffentlicht: (2025)
von: Ranjbar, Ali, et al.
Veröffentlicht: (2025)
A Novel FPGA-based CNN Hardware Accelerator: Optimization for Convolutional Layers using Karatsuba Ofman Multiplier
von: Sarkar, Amit
Veröffentlicht: (2024)
von: Sarkar, Amit
Veröffentlicht: (2024)
Ähnliche Einträge
-
Real Time FPGA Based CNNs for Detection, Classification, and Tracking in Autonomous Systems: State of the Art Designs and Optimizations
von: Sali, Safa Mohammed, et al.
Veröffentlicht: (2025) -
Edge GPU Aware Multiple AI Model Pipeline for Accelerated MRI Reconstruction and Analysis
von: Majeed, Ashiyana Abdul, et al.
Veröffentlicht: (2025) -
Real-time Object Detection and Associated Hardware Accelerators Targeting Autonomous Vehicles: A Review
von: Sali, Safa, et al.
Veröffentlicht: (2025) -
Scheduling Techniques of AI Models on Modern Heterogeneous Edge GPU -- A Critical Review
von: Majeed, Ashiyana Abdul, et al.
Veröffentlicht: (2025) -
Hardware Acceleration in Portable MRIs: State of the Art and Future Prospects
von: Habsi, Omar Al, et al.
Veröffentlicht: (2025)