A Survey on Design Methodologies for Accelerating Deep Learning on Heterogeneous Architectures
Fuente:
arXiv
Saved in:
| Main Authors: | Curzel, Serena, Ferrandi, Fabrizio, Fiorin, Leandro, Ielmini, Daniele, Silvano, Cristina, Conti, Francesco, Bompani, Luca, Benini, Luca, Calore, Enrico, Schifano, Sebastiano Fabio, Zambelli, Cristian, Palesi, Maurizio, Ascia, Giuseppe, Russo, Enrico, Cardellini, Valeria, Filippone, Salvatore, Presti, Francesco Lo, Perri, Stefania |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Survey on Deep Learning Hardware Accelerators for Heterogeneous HPC Platforms
by: Silvano, Cristina, et al.
Published: (2023)
by: Silvano, Cristina, et al.
Published: (2023)
CHAOS: Controlled Hardware fAult injectOr System for gem5
by: Vinciguerra, Elio, et al.
Published: (2026)
by: Vinciguerra, Elio, et al.
Published: (2026)
Attention-Based Deep Reinforcement Learning for Qubit Allocation in Modular Quantum Architectures
by: Russo, Enrico, et al.
Published: (2024)
by: Russo, Enrico, et al.
Published: (2024)
Assessing the Role of Communication in Scalable Multi-Core Quantum Architectures
by: Palesi, Maurizio, et al.
Published: (2024)
by: Palesi, Maurizio, et al.
Published: (2024)
Deep Reinforcement Learning based Online Scheduling Policy for Deep Neural Network Multi-Tenant Multi-Accelerator Systems
by: Blanco, Francesco G., et al.
Published: (2024)
by: Blanco, Francesco G., et al.
Published: (2024)
Towards Fair and Firm Real-Time Scheduling in DNN Multi-Tenant Multi-Accelerator Systems via Reinforcement Learning
by: Russo, Enrico, et al.
Published: (2024)
by: Russo, Enrico, et al.
Published: (2024)
MR2-ByteTrack: CNN and Transformer-based Video Object Detection for AI-augmented Embedded Vision Sensor Nodes
by: Bompani, Luca, et al.
Published: (2026)
by: Bompani, Luca, et al.
Published: (2026)
Multi-resolution Rescored ByteTrack for Video Object Detection on Ultra-low-power Embedded Systems
by: Bompani, Luca, et al.
Published: (2024)
by: Bompani, Luca, et al.
Published: (2024)
TeleSABRE: Layout Synthesis in Multi-Core Quantum Systems with Teleport Interconnect
by: Russo, Enrico, et al.
Published: (2025)
by: Russo, Enrico, et al.
Published: (2025)
Multi-Objective Hardware-Mapping Co-Optimisation for Multi-DNN Workloads on Chiplet-based Accelerators
by: Das, Abhijit, et al.
Published: (2022)
by: Das, Abhijit, et al.
Published: (2022)
MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs
by: Russo, Enrico, et al.
Published: (2026)
by: Russo, Enrico, et al.
Published: (2026)
Instruction-Directed MAC for Efficient Classical Communication in Scalable Multi-Chip Quantum Systems
by: Palesi, Maurizio, et al.
Published: (2025)
by: Palesi, Maurizio, et al.
Published: (2025)
Accelerating Image-based Pest Detection on a Heterogeneous Multi-core Microcontroller
by: Bompani, Luca, et al.
Published: (2024)
by: Bompani, Luca, et al.
Published: (2024)
Deeploy: Enabling Energy-Efficient Deployment of Small Language Models On Heterogeneous Microcontrollers
by: Scherer, Moritz, et al.
Published: (2024)
by: Scherer, Moritz, et al.
Published: (2024)
Temporal Numeric Planning with Patterns
by: Cardellini, Matteo, et al.
Published: (2024)
by: Cardellini, Matteo, et al.
Published: (2024)
Symbolic Pattern Temporal Numeric Planning with Intermediate Conditions and Effects
by: Cardellini, Matteo, et al.
Published: (2026)
by: Cardellini, Matteo, et al.
Published: (2026)
Distilling Tiny and Ultra-fast Deep Neural Networks for Autonomous Navigation on Nano-UAVs
by: Lamberti, Lorenzo, et al.
Published: (2024)
by: Lamberti, Lorenzo, et al.
Published: (2024)
Combining Local and Global Perception for Autonomous Navigation on Nano-UAVs
by: Lamberti, Lorenzo, et al.
Published: (2024)
by: Lamberti, Lorenzo, et al.
Published: (2024)
A Data-Driven Approach to Dataflow-Aware Online Scheduling for Graph Neural Network Inference
by: Puigdemont, Pol, et al.
Published: (2024)
by: Puigdemont, Pol, et al.
Published: (2024)
¿Quién habla en la oreja de Einstein? Arte indígena contemporáneo en el estado de Chiapas (México)
by: Luca D'Ascia
Published: (2007)
by: Luca D'Ascia
Published: (2007)
Evaluating IOMMU-Based Shared Virtual Addressing for RISC-V Embedded Heterogeneous SoCs
by: Koenig, Cyril, et al.
Published: (2025)
by: Koenig, Cyril, et al.
Published: (2025)
Symbolic Numeric Planning with Patterns
by: Cardellini, Matteo, et al.
Published: (2023)
by: Cardellini, Matteo, et al.
Published: (2023)
Assessing the Role of Communication in Modular Multi-Core Quantum Systems
by: Palesi, Maurizio, et al.
Published: (2025)
by: Palesi, Maurizio, et al.
Published: (2025)
Characterizing Information Shared by Participants to Coding Challenges: The Case of Advent of Code
by: Cauteruccio, Francesco, et al.
Published: (2024)
by: Cauteruccio, Francesco, et al.
Published: (2024)
Open-Source Heterogeneous SoCs for AI: The PULP Platform Experience
by: Conti, Francesco, et al.
Published: (2024)
by: Conti, Francesco, et al.
Published: (2024)
Fused-Tiled Layers: Minimizing Data Movement on RISC-V SoCs with Software-Managed Caches
by: Jung, Victor J. B., et al.
Published: (2025)
by: Jung, Victor J. B., et al.
Published: (2025)
Hybrid Modular Redundancy: Exploring Modular Redundancy Approaches in RISC-V Multi-Core Computing Clusters for Reliable Processing in Space
by: Rogenmoser, Michael, et al.
Published: (2023)
by: Rogenmoser, Michael, et al.
Published: (2023)
MXDOTP: A RISC-V ISA Extension for Enabling Microscaling (MX) Floating-Point Dot Products
by: İslamoğlu, Gamze, et al.
Published: (2025)
by: İslamoğlu, Gamze, et al.
Published: (2025)
DataPix4: A C++ framework for Timepix4 configuration and read-out
by: Cavallini, Viola, et al.
Published: (2025)
by: Cavallini, Viola, et al.
Published: (2025)
Lightweight Software Kernels and Hardware Extensions for Efficient Sparse Deep Neural Networks on Microcontrollers
by: Daghero, Francesco, et al.
Published: (2025)
by: Daghero, Francesco, et al.
Published: (2025)
Fredholm alternative for a general class of nonlocal operators
by: De Pas, Francesco, et al.
Published: (2026)
by: De Pas, Francesco, et al.
Published: (2026)
Reconstructing double-well potentials from transition layers in long-range phase coexistence models
by: Dipierro, Serena, et al.
Published: (2026)
by: Dipierro, Serena, et al.
Published: (2026)
Optimal decay of heteroclinic solutions of the fractional Allen-Cahn equation with a degenerate potential
by: De Pas, Francesco, et al.
Published: (2026)
by: De Pas, Francesco, et al.
Published: (2026)
Long-range phase coexistence models with degenerate potentials
by: De Pas, Francesco, et al.
Published: (2026)
by: De Pas, Francesco, et al.
Published: (2026)
Work-In-Progress: Accelerating Numpy With OpenBLAS For Open-Source RISC-V Chips
by: Koenig, Cyril, et al.
Published: (2025)
by: Koenig, Cyril, et al.
Published: (2025)
Non-unital noise in a superconducting quantum computer as a computational resource for reservoir computing
by: Monzani, Francesco, et al.
Published: (2024)
by: Monzani, Francesco, et al.
Published: (2024)
Quantum reservoir computing induced by controllable damping
by: Ricci, Emanuele, et al.
Published: (2025)
by: Ricci, Emanuele, et al.
Published: (2025)
Designing Control Barrier Function via Probabilistic Enumeration for Safe Reinforcement Learning Navigation
by: Marzari, Luca, et al.
Published: (2025)
by: Marzari, Luca, et al.
Published: (2025)
Advancing risk stratification in pulmonary arterial hypertension through echocardiographic innovation
by: Angelica Cersosimo, et al.
Published: (2024)
by: Angelica Cersosimo, et al.
Published: (2024)
A new approach to rating scale definition with quantum-inspired optimization
by: Spada, Patrizio, et al.
Published: (2026)
by: Spada, Patrizio, et al.
Published: (2026)
Similar Items
-
A Survey on Deep Learning Hardware Accelerators for Heterogeneous HPC Platforms
by: Silvano, Cristina, et al.
Published: (2023) -
CHAOS: Controlled Hardware fAult injectOr System for gem5
by: Vinciguerra, Elio, et al.
Published: (2026) -
Attention-Based Deep Reinforcement Learning for Qubit Allocation in Modular Quantum Architectures
by: Russo, Enrico, et al.
Published: (2024) -
Assessing the Role of Communication in Scalable Multi-Core Quantum Architectures
by: Palesi, Maurizio, et al.
Published: (2024) -
Deep Reinforcement Learning based Online Scheduling Policy for Deep Neural Network Multi-Tenant Multi-Accelerator Systems
by: Blanco, Francesco G., et al.
Published: (2024)