Designing DNNs for a trade-off between robustness and processing performance in embedded devices

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Gutiérrez-Zaballa, Jon, Basterretxea, Koldo, Echanobe, Javier
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909416860680192
author Gutiérrez-Zaballa, Jon
Basterretxea, Koldo
Echanobe, Javier
author_facet Gutiérrez-Zaballa, Jon
Basterretxea, Koldo
Echanobe, Javier
contents Machine learning-based embedded systems employed in safety-critical applications such as aerospace and autonomous driving need to be robust against perturbations produced by soft errors. Soft errors are an increasing concern in modern digital processors since smaller transistor geometries and lower voltages give electronic devices a higher sensitivity to background radiation. The resilience of deep neural network (DNN) models to perturbations in their parameters is determined, to a large extent, by the structure of the model itself, and also by the selected numerical representation and used arithmetic precision. When compression techniques such as model pruning and model quantization are applied to reduce memory footprint and computational complexity for deployment, both model structure and numerical representation are modified and thus, soft error robustness also changes. In this sense, although the choice of activation functions (AFs) in DNN models is frequently ignored, it conditions not only their accuracy and trainability, but also compressibility rates and numerical robustness. This paper investigates the suitability of using bounded AFs to improve model robustness against DNN parameter perturbations, assessing at the same time the impact of this choice on deployment in terms of model accuracy, compressibility, and computational burden. In particular, we analyze encoder-decoder fully convolutional models aimed at performing semantic segmentation tasks on hyperspectral images for scene understanding in autonomous driving. Deployment characterization is performed experimentally on an AMD-Xilinx's KV260 SoM.
format Preprint
id arxiv_https___arxiv_org_abs_2412_03682
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Designing DNNs for a trade-off between robustness and processing performance in embedded devices
Gutiérrez-Zaballa, Jon
Basterretxea, Koldo
Echanobe, Javier
Machine Learning
Artificial Intelligence
Hardware Architecture
Computer Vision and Pattern Recognition
Image and Video Processing
Machine learning-based embedded systems employed in safety-critical applications such as aerospace and autonomous driving need to be robust against perturbations produced by soft errors. Soft errors are an increasing concern in modern digital processors since smaller transistor geometries and lower voltages give electronic devices a higher sensitivity to background radiation. The resilience of deep neural network (DNN) models to perturbations in their parameters is determined, to a large extent, by the structure of the model itself, and also by the selected numerical representation and used arithmetic precision. When compression techniques such as model pruning and model quantization are applied to reduce memory footprint and computational complexity for deployment, both model structure and numerical representation are modified and thus, soft error robustness also changes. In this sense, although the choice of activation functions (AFs) in DNN models is frequently ignored, it conditions not only their accuracy and trainability, but also compressibility rates and numerical robustness. This paper investigates the suitability of using bounded AFs to improve model robustness against DNN parameter perturbations, assessing at the same time the impact of this choice on deployment in terms of model accuracy, compressibility, and computational burden. In particular, we analyze encoder-decoder fully convolutional models aimed at performing semantic segmentation tasks on hyperspectral images for scene understanding in autonomous driving. Deployment characterization is performed experimentally on an AMD-Xilinx's KV260 SoM.
title Designing DNNs for a trade-off between robustness and processing performance in embedded devices
topic Machine Learning
Artificial Intelligence
Hardware Architecture
Computer Vision and Pattern Recognition
Image and Video Processing
url https://arxiv.org/abs/2412.03682