Benchmarking Deep Learning Models on NVIDIA Jetson Nano for Real-Time Systems: An Empirical Investigation
Fuente:
arXiv
Guardado en:
| Autores principales: | Swaminathan, Tushar Prasanna, Silver, Christopher, Akilan, Thangarajah |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Primitive-Driven Acceleration of Hyperdimensional Computing for Real-Time Image Classification
por: Parikh, Dhruv, et al.
Publicado: (2026)
por: Parikh, Dhruv, et al.
Publicado: (2026)
On-Orbit Real-Time Wildfire Detection Under On-Board Constraints
por: Rötzer, Matthias, et al.
Publicado: (2026)
por: Rötzer, Matthias, et al.
Publicado: (2026)
Real-Time Object Detection and Classification using YOLO for Edge FPGAs
por: Amin, Rashed Al, et al.
Publicado: (2025)
por: Amin, Rashed Al, et al.
Publicado: (2025)
Neo: Real-Time On-Device 3D Gaussian Splatting with Reuse-and-Update Sorting Acceleration
por: Oh, Changhun, et al.
Publicado: (2025)
por: Oh, Changhun, et al.
Publicado: (2025)
hARMS: A Hardware Acceleration Architecture for Real-Time Event-Based Optical Flow
por: Stumpp, Daniel C., et al.
Publicado: (2021)
por: Stumpp, Daniel C., et al.
Publicado: (2021)
AppSign: Multi-level Approximate Computing for Real-Time Traffic Sign Recognition in Autonomous Vehicles
por: Omidian, Fatemeh, et al.
Publicado: (2024)
por: Omidian, Fatemeh, et al.
Publicado: (2024)
A Parameterizable Convolution Accelerator for Embedded Deep Learning Applications
por: Mousouliotis, Panagiotis, et al.
Publicado: (2026)
por: Mousouliotis, Panagiotis, et al.
Publicado: (2026)
Thermal Imaging-based Real-time Fall Detection using Motion Flow and Attention-enhanced Convolutional Recurrent Architecture
por: Silver, Christopher, et al.
Publicado: (2025)
por: Silver, Christopher, et al.
Publicado: (2025)
Smaller, Faster, Cheaper: Architectural Designs for Efficient Machine Learning
por: Walton, Steven
Publicado: (2025)
por: Walton, Steven
Publicado: (2025)
Profiling Concurrent Vision Inference Workloads on NVIDIA Jetson -- Extended
por: Chakraborty, Abhinaba, et al.
Publicado: (2025)
por: Chakraborty, Abhinaba, et al.
Publicado: (2025)
Ditto: Accelerating Diffusion Model via Temporal Value Similarity
por: Kim, Sungbin, et al.
Publicado: (2025)
por: Kim, Sungbin, et al.
Publicado: (2025)
CRISP: Hybrid Structured Sparsity for Class-aware Model Pruning
por: Aggarwal, Shivam, et al.
Publicado: (2023)
por: Aggarwal, Shivam, et al.
Publicado: (2023)
NeuralFuse: Learning to Recover the Accuracy of Access-Limited Neural Network Inference in Low-Voltage Regimes
por: Sun, Hao-Lun, et al.
Publicado: (2023)
por: Sun, Hao-Lun, et al.
Publicado: (2023)
Uni-Render: A Unified Accelerator for Real-Time Rendering Across Diverse Neural Renderers
por: Li, Chaojian, et al.
Publicado: (2025)
por: Li, Chaojian, et al.
Publicado: (2025)
AHCQ-SAM: Toward Accurate and Hardware-Compatible Post-Training Segment Anything Model Quantization
por: Zhang, Wenlun, et al.
Publicado: (2025)
por: Zhang, Wenlun, et al.
Publicado: (2025)
ViM-Q: Scalable Algorithm-Hardware Co-Design for Vision Mamba Model Inference on FPGA
por: Lyu, Shengzhe, et al.
Publicado: (2026)
por: Lyu, Shengzhe, et al.
Publicado: (2026)
Real-Time Semantic Segmentation of Aerial Images Using an Embedded U-Net: A Comparison of CPU, GPU, and FPGA Workflows
por: Posso, Julien, et al.
Publicado: (2025)
por: Posso, Julien, et al.
Publicado: (2025)
SteROI-D: System Design and Mapping for Stereo Depth Inference on Regions of Interest
por: Erhardt, Jack, et al.
Publicado: (2025)
por: Erhardt, Jack, et al.
Publicado: (2025)
Energy Efficient Exact and Approximate Systolic Array Architecture for Matrix Multiplication
por: Jaswal, Pragun, et al.
Publicado: (2025)
por: Jaswal, Pragun, et al.
Publicado: (2025)
Stella Nera: A Differentiable Maddness-Based Hardware Accelerator for Efficient Approximate Matrix Multiplication
por: Schönleber, Jannis, et al.
Publicado: (2023)
por: Schönleber, Jannis, et al.
Publicado: (2023)
SMOF: Streaming Modern CNNs on FPGAs with Smart Off-Chip Eviction
por: Toupas, Petros, et al.
Publicado: (2024)
por: Toupas, Petros, et al.
Publicado: (2024)
Performance Analysis of Edge and In-Sensor AI Processors: A Comparative Review
por: Capogrosso, Luigi, et al.
Publicado: (2026)
por: Capogrosso, Luigi, et al.
Publicado: (2026)
QUILL: An Algorithm-Architecture Co-Design for Cache-Local Deformable Attention
por: Oh, Hyunwoo, et al.
Publicado: (2025)
por: Oh, Hyunwoo, et al.
Publicado: (2025)
Neuro-Channel Networks: A Multiplication-Free Architecture by Biological Signal Transmission
por: Mete, Emrah, et al.
Publicado: (2026)
por: Mete, Emrah, et al.
Publicado: (2026)
ViTCoD: Vision Transformer Acceleration via Dedicated Algorithm and Accelerator Co-Design
por: You, Haoran, et al.
Publicado: (2022)
por: You, Haoran, et al.
Publicado: (2022)
HARFLOW3D: A Latency-Oriented 3D-CNN Accelerator Toolflow for HAR on FPGA Devices
por: Toupas, Petros, et al.
Publicado: (2023)
por: Toupas, Petros, et al.
Publicado: (2023)
Mix-and-Match Pruning: Globally Guided Layer-Wise Sparsification of DNNs
por: Monachan, Danial, et al.
Publicado: (2026)
por: Monachan, Danial, et al.
Publicado: (2026)
CAMO: Correlation-Aware Mask Optimization with Modulated Reinforcement Learning
por: Liang, Xiaoxiao, et al.
Publicado: (2024)
por: Liang, Xiaoxiao, et al.
Publicado: (2024)
Co-designing a Sub-millisecond Latency Event-based Eye Tracking System with Submanifold Sparse CNN
por: Zhang, Baoheng, et al.
Publicado: (2024)
por: Zhang, Baoheng, et al.
Publicado: (2024)
Investigating Robot Dogs for Construction Monitoring: A Comparative Analysis of Specifications and On-site Requirements
por: Torres, Miguel Arturo Vega, et al.
Publicado: (2024)
por: Torres, Miguel Arturo Vega, et al.
Publicado: (2024)
Neural Architecture Search of Hybrid Models for NPU-CIM Heterogeneous AR/VR Devices
por: Zhao, Yiwei, et al.
Publicado: (2024)
por: Zhao, Yiwei, et al.
Publicado: (2024)
Vision Transformers on the Edge: A Comprehensive Survey of Model Compression and Acceleration Strategies
por: Saha, Shaibal, et al.
Publicado: (2025)
por: Saha, Shaibal, et al.
Publicado: (2025)
SF-MMCN: Low-Power Sever Flow Multi-Mode Diffusion Model Accelerator
por: Hsu, Huan-Ke, et al.
Publicado: (2024)
por: Hsu, Huan-Ke, et al.
Publicado: (2024)
When Spike Sparsity Does Not Translate to Deployed Cost: VS-WNO on Jetson Orin Nano
por: Yoo, Jason, et al.
Publicado: (2026)
por: Yoo, Jason, et al.
Publicado: (2026)
Real-World Deployment of a Lane Change Prediction Architecture Based on Knowledge Graph Embeddings and Bayesian Inference
por: Manzour, M., et al.
Publicado: (2025)
por: Manzour, M., et al.
Publicado: (2025)
Real-Time Spacecraft Pose Estimation Using Mixed-Precision Quantized Neural Network on COTS Reconfigurable MPSoC
por: Posso, Julien, et al.
Publicado: (2024)
por: Posso, Julien, et al.
Publicado: (2024)
Accelerating 3D Gaussian Splatting with Neural Sorting and Axis-Oriented Rasterization
por: Wang, Zhican, et al.
Publicado: (2025)
por: Wang, Zhican, et al.
Publicado: (2025)
RaGNNarok: A Light-Weight Graph Neural Network for Enhancing Radar Point Clouds on Unmanned Ground Vehicles
por: Hunt, David, et al.
Publicado: (2025)
por: Hunt, David, et al.
Publicado: (2025)
On Latency Predictors for Neural Architecture Search
por: Akhauri, Yash, et al.
Publicado: (2024)
por: Akhauri, Yash, et al.
Publicado: (2024)
Evolving Layer-Specific Scalar Functions for Hardware-Aware Transformer Adaptation
por: Carrigg, Kieran, et al.
Publicado: (2026)
por: Carrigg, Kieran, et al.
Publicado: (2026)
Ejemplares similares
-
Primitive-Driven Acceleration of Hyperdimensional Computing for Real-Time Image Classification
por: Parikh, Dhruv, et al.
Publicado: (2026) -
On-Orbit Real-Time Wildfire Detection Under On-Board Constraints
por: Rötzer, Matthias, et al.
Publicado: (2026) -
Real-Time Object Detection and Classification using YOLO for Edge FPGAs
por: Amin, Rashed Al, et al.
Publicado: (2025) -
Neo: Real-Time On-Device 3D Gaussian Splatting with Reuse-and-Update Sorting Acceleration
por: Oh, Changhun, et al.
Publicado: (2025) -
hARMS: A Hardware Acceleration Architecture for Real-Time Event-Based Optical Flow
por: Stumpp, Daniel C., et al.
Publicado: (2021)