MATCH: Model-Aware TVM-based Compilation for Heterogeneous Edge Devices
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hamdi, Mohamed Amine, Daghero, Francesco, Sarda, Giuseppe Maria, Van Delm, Josse, Symons, Arne, Benini, Luca, Verhelst, Marian, Pagliari, Daniele Jahier, Burrello, Alessio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HTVM: Efficient Neural Network Deployment On Heterogeneous TinyML Platforms
von: Van Delm, Josse, et al.
Veröffentlicht: (2024)
von: Van Delm, Josse, et al.
Veröffentlicht: (2024)
Accelerating Depthwise Separable Convolutions on Ultra-Low-Power Devices
von: Daghero, Francesco, et al.
Veröffentlicht: (2024)
von: Daghero, Francesco, et al.
Veröffentlicht: (2024)
Lightweight Software Kernels and Hardware Extensions for Efficient Sparse Deep Neural Networks on Microcontrollers
von: Daghero, Francesco, et al.
Veröffentlicht: (2025)
von: Daghero, Francesco, et al.
Veröffentlicht: (2025)
Optimizing Foundation Model Inference on a Many-tiny-core Open-source RISC-V Platform
von: Potocnik, Viviane, et al.
Veröffentlicht: (2024)
von: Potocnik, Viviane, et al.
Veröffentlicht: (2024)
MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs
von: Russo, Enrico, et al.
Veröffentlicht: (2026)
von: Russo, Enrico, et al.
Veröffentlicht: (2026)
Integrating SystemC-AMS Power Modeling with a RISC-V ISS for Virtual Prototyping of Battery-operated Embedded Devices
von: Hamdi, Mohamed Amine, et al.
Veröffentlicht: (2024)
von: Hamdi, Mohamed Amine, et al.
Veröffentlicht: (2024)
Foundation Models for Structural Health Monitoring
von: Benfenati, Luca, et al.
Veröffentlicht: (2024)
von: Benfenati, Luca, et al.
Veröffentlicht: (2024)
V-Seek: Accelerating LLM Reasoning on Open-hardware Server-class RISC-V Platforms
von: Rodrigo, Javier J. Poveda, et al.
Veröffentlicht: (2025)
von: Rodrigo, Javier J. Poveda, et al.
Veröffentlicht: (2025)
OnDA: On-device Channel Pruning for Efficient Personalized Keyword Spotting
von: Risso, Matteo, et al.
Veröffentlicht: (2026)
von: Risso, Matteo, et al.
Veröffentlicht: (2026)
Optimizing DNN Inference on Multi-Accelerator SoCs at Training-time
von: Risso, Matteo, et al.
Veröffentlicht: (2024)
von: Risso, Matteo, et al.
Veröffentlicht: (2024)
BlankSkip: Early-exit Object Detection onboard Nano-drones
von: Marra, Carlo, et al.
Veröffentlicht: (2026)
von: Marra, Carlo, et al.
Veröffentlicht: (2026)
BISeizuRe: BERT-Inspired Seizure Data Representation to Improve Epilepsy Monitoring
von: Benfenati, Luca, et al.
Veröffentlicht: (2024)
von: Benfenati, Luca, et al.
Veröffentlicht: (2024)
A Multi-level Compiler Backend for Accelerated Micro-kernels Targeting RISC-V ISA Extensions
von: Lopoukhine, Alexandre, et al.
Veröffentlicht: (2025)
von: Lopoukhine, Alexandre, et al.
Veröffentlicht: (2025)
Optimized Deployment of Deep Neural Networks for Visual Pose Estimation on Nano-drones
von: Risso, Matteo, et al.
Veröffentlicht: (2024)
von: Risso, Matteo, et al.
Veröffentlicht: (2024)
Improving Continual Learning for Gaussian Splatting based Environments Reconstruction on Commercial Off-the-Shelf Edge Devices
von: Zaino, Ivan, et al.
Veröffentlicht: (2026)
von: Zaino, Ivan, et al.
Veröffentlicht: (2026)
SilentWear: an Ultra-Low Power Wearable System for EMG-based Silent Speech Recognition
von: Spacone, Giusy, et al.
Veröffentlicht: (2026)
von: Spacone, Giusy, et al.
Veröffentlicht: (2026)
Optimization and Deployment of Deep Neural Networks for PPG-based Blood Pressure Estimation Targeting Low-power Wearables
von: Burrello, Alessio, et al.
Veröffentlicht: (2024)
von: Burrello, Alessio, et al.
Veröffentlicht: (2024)
Before Parc Fermé: RL-Time Pruning for Efficient Embodied LLMs in Autonomous Driving
von: Benfenati, Luca, et al.
Veröffentlicht: (2026)
von: Benfenati, Luca, et al.
Veröffentlicht: (2026)
Performance evaluation of acceleration of convolutional layers on OpenEdgeCGRA
von: Carpentieri, Nicolò, et al.
Veröffentlicht: (2024)
von: Carpentieri, Nicolò, et al.
Veröffentlicht: (2024)
Fine-Grained Fusion: The Missing Piece in Area-Efficient State Space Model Acceleration
von: Geens, Robin, et al.
Veröffentlicht: (2025)
von: Geens, Robin, et al.
Veröffentlicht: (2025)
SALSA: Simulated Annealing based Loop-Ordering Scheduler for DNN Accelerators
von: Jung, Victor J. B., et al.
Veröffentlicht: (2023)
von: Jung, Victor J. B., et al.
Veröffentlicht: (2023)
HW-SW Optimization of DNNs for Privacy-preserving People Counting on Low-resolution Infrared Arrays
von: Risso, Matteo, et al.
Veröffentlicht: (2024)
von: Risso, Matteo, et al.
Veröffentlicht: (2024)
EnhancePPG: Improving PPG-based Heart Rate Estimation with Self-Supervision and Augmentation
von: Benfenati, Luca, et al.
Veröffentlicht: (2024)
von: Benfenati, Luca, et al.
Veröffentlicht: (2024)
End-to-end Automated Deep Neural Network Optimization for PPG-based Blood Pressure Estimation on Wearables
von: Carlucci, Francesco, et al.
Veröffentlicht: (2026)
von: Carlucci, Francesco, et al.
Veröffentlicht: (2026)
Coupling Neural Networks and Physics Equations For Li-Ion Battery State-of-Charge Prediction
von: Pollo, Giovanni, et al.
Veröffentlicht: (2024)
von: Pollo, Giovanni, et al.
Veröffentlicht: (2024)
Hardware-Algorithm Co-Optimization of Early-Exit Neural Networks for Multi-Core Edge Accelerators
von: Zniber, Alaa, et al.
Veröffentlicht: (2025)
von: Zniber, Alaa, et al.
Veröffentlicht: (2025)
Late Breaking Results: CHESSY: Coupled Hybrid Emulation with SystemC-FPGA Synchronization
von: Ruotolo, Lorenzo, et al.
Veröffentlicht: (2026)
von: Ruotolo, Lorenzo, et al.
Veröffentlicht: (2026)
Joint Pruning and Channel-wise Mixed-Precision Quantization for Efficient Deep Neural Networks
von: Motetti, Beatrice Alessandra, et al.
Veröffentlicht: (2024)
von: Motetti, Beatrice Alessandra, et al.
Veröffentlicht: (2024)
DeFiNES: Enabling Fast Exploration of the Depth-first Scheduling Space for DNN Accelerators through Analytical Modeling
von: Mei, Linyan, et al.
Veröffentlicht: (2022)
von: Mei, Linyan, et al.
Veröffentlicht: (2022)
How to keep pushing ML accelerator performance? Know your rooflines!
von: Verhelst, Marian, et al.
Veröffentlicht: (2025)
von: Verhelst, Marian, et al.
Veröffentlicht: (2025)
Stream: Design Space Exploration of Layer-Fused DNNs on Heterogeneous Dataflow Accelerators
von: Symons, Arne, et al.
Veröffentlicht: (2022)
von: Symons, Arne, et al.
Veröffentlicht: (2022)
Adaptive Deep Learning for Efficient Visual Pose Estimation aboard Ultra-low-power Nano-drones
von: Motetti, Beatrice Alessandra, et al.
Veröffentlicht: (2024)
von: Motetti, Beatrice Alessandra, et al.
Veröffentlicht: (2024)
OpenGeMM: A High-Utilization GeMM Accelerator Generator with Lightweight RISC-V Control and Tight Memory Coupling
von: Yi, Xiaoling, et al.
Veröffentlicht: (2024)
von: Yi, Xiaoling, et al.
Veröffentlicht: (2024)
Don't be so Stief! Learning KV Cache low-rank approximation over the Stiefel manifold
von: Benfenati, Luca, et al.
Veröffentlicht: (2026)
von: Benfenati, Luca, et al.
Veröffentlicht: (2026)
Multi-modal On-Device Learning for Monocular Depth Estimation on Ultra-low-power MCUs
von: Nadalini, Davide, et al.
Veröffentlicht: (2025)
von: Nadalini, Davide, et al.
Veröffentlicht: (2025)
Deep Recommender Models Inference: Automatic Asymmetric Data Flow Optimization
von: Ruggeri, Giuseppe, et al.
Veröffentlicht: (2025)
von: Ruggeri, Giuseppe, et al.
Veröffentlicht: (2025)
An Open-Source HW-SW Co-Development Framework Enabling Efficient Multi-Accelerator Systems
von: Antonio, Ryan Albert, et al.
Veröffentlicht: (2025)
von: Antonio, Ryan Albert, et al.
Veröffentlicht: (2025)
MEbots: Integrating a RISC-V Virtual Platform with a Robotic Simulator for Energy-aware Design
von: Pollo, Giovanni, et al.
Veröffentlicht: (2025)
von: Pollo, Giovanni, et al.
Veröffentlicht: (2025)
Optimizing Layer-Fused Scheduling of Transformer Networks on Multi-accelerator Platforms
von: Colleman, Steven, et al.
Veröffentlicht: (2024)
von: Colleman, Steven, et al.
Veröffentlicht: (2024)
MONET: Modeling and Optimization of neural NEtwork Training from Edge to Data Centers
von: Morlier, Jérémy, et al.
Veröffentlicht: (2026)
von: Morlier, Jérémy, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
HTVM: Efficient Neural Network Deployment On Heterogeneous TinyML Platforms
von: Van Delm, Josse, et al.
Veröffentlicht: (2024) -
Accelerating Depthwise Separable Convolutions on Ultra-Low-Power Devices
von: Daghero, Francesco, et al.
Veröffentlicht: (2024) -
Lightweight Software Kernels and Hardware Extensions for Efficient Sparse Deep Neural Networks on Microcontrollers
von: Daghero, Francesco, et al.
Veröffentlicht: (2025) -
Optimizing Foundation Model Inference on a Many-tiny-core Open-source RISC-V Platform
von: Potocnik, Viviane, et al.
Veröffentlicht: (2024) -
MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs
von: Russo, Enrico, et al.
Veröffentlicht: (2026)