Tiny Deep Ensemble: Uncertainty Estimation in Edge AI Accelerators via Ensembling Normalization Layers with Shared Weights
Fuente:
arXiv
Salvato in:
| Autori principali: | Ahmed, Soyed Tuhin, Hefenbrock, Michael, Tahoori, Mehdi B. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Minimizing CGYRO HPC Communication Costs in Ensembles with XGYRO by Sharing the Collisional Constant Tensor Structure
di: Sfiligoi, Igor, et al.
Pubblicazione: (2025)
di: Sfiligoi, Igor, et al.
Pubblicazione: (2025)
Accelerated Digital Twin Learning for Edge AI: A Comparison of FPGA and Mobile GPU
di: Xu, Bin, et al.
Pubblicazione: (2025)
di: Xu, Bin, et al.
Pubblicazione: (2025)
Synergy: Towards On-Body AI via Tiny AI Accelerator Collaboration on Wearables
di: Gong, Taesik, et al.
Pubblicazione: (2023)
di: Gong, Taesik, et al.
Pubblicazione: (2023)
On the Impact of White-box Deployment Strategies for Edge AI on Latency and Model Performance
di: Singh, Jaskirat, et al.
Pubblicazione: (2024)
di: Singh, Jaskirat, et al.
Pubblicazione: (2024)
Ensemble Method for System Failure Detection Using Large-Scale Telemetry Data
di: Mudgal, Priyanka, et al.
Pubblicazione: (2024)
di: Mudgal, Priyanka, et al.
Pubblicazione: (2024)
Hardware Utilization and Inference Performance of Edge Object Detection Under Fault Injection
di: Pasandideh, Faezeh, et al.
Pubblicazione: (2026)
di: Pasandideh, Faezeh, et al.
Pubblicazione: (2026)
An Ensemble Scheme for Proactive Dominant Data Migration of Pervasive Tasks at the Edge
di: Boulougaris, Georgios, et al.
Pubblicazione: (2024)
di: Boulougaris, Georgios, et al.
Pubblicazione: (2024)
Characterizing the Performance of Accelerated Jetson Edge Devices for Training Deep Learning Models
di: K., Prashanthi S., et al.
Pubblicazione: (2025)
di: K., Prashanthi S., et al.
Pubblicazione: (2025)
Quality Scalable Quantization Methodology for Deep Learning on Edge
di: Khaliq, Salman Abdul, et al.
Pubblicazione: (2024)
di: Khaliq, Salman Abdul, et al.
Pubblicazione: (2024)
E-QUARTIC: Energy Efficient Edge Ensemble of Convolutional Neural Networks for Resource-Optimized Learning
di: Zhang, Le, et al.
Pubblicazione: (2024)
di: Zhang, Le, et al.
Pubblicazione: (2024)
Portable, heterogeneous ensemble workflows at scale using libEnsemble
di: Hudson, Stephen, et al.
Pubblicazione: (2024)
di: Hudson, Stephen, et al.
Pubblicazione: (2024)
TinyServe: Query-Aware Cache Selection for Efficient LLM Serving
di: Liu, Dong, et al.
Pubblicazione: (2025)
di: Liu, Dong, et al.
Pubblicazione: (2025)
Few-Shot Testing: Estimating Uncertainty of Memristive Deep Neural Networks Using One Bayesian Test Vector
di: Ahmed, Soyed Tuhin, et al.
Pubblicazione: (2024)
di: Ahmed, Soyed Tuhin, et al.
Pubblicazione: (2024)
Efficient Federated Finetuning of Tiny Transformers with Resource-Constrained Devices
di: Pfeiffer, Kilian, et al.
Pubblicazione: (2024)
di: Pfeiffer, Kilian, et al.
Pubblicazione: (2024)
Efficient Edge AI: Deploying Convolutional Neural Networks on FPGA with the Gemmini Accelerator
di: Peccia, Federico Nicolas, et al.
Pubblicazione: (2024)
di: Peccia, Federico Nicolas, et al.
Pubblicazione: (2024)
Adaptive AI-based Decentralized Resource Management in the Cloud-Edge Continuum
di: Li, Lanpei, et al.
Pubblicazione: (2025)
di: Li, Lanpei, et al.
Pubblicazione: (2025)
Dora: QoE-Aware Hybrid Parallelism for Distributed Edge AI
di: Jin, Jianli, et al.
Pubblicazione: (2025)
di: Jin, Jianli, et al.
Pubblicazione: (2025)
LogAct: Enabling Agentic Reliability via Shared Logs
di: Balakrishnan, Mahesh, et al.
Pubblicazione: (2026)
di: Balakrishnan, Mahesh, et al.
Pubblicazione: (2026)
Pico-Cloud: Cloud Infrastructure for Tiny Edge Devices
di: Guri, Mordechai
Pubblicazione: (2025)
di: Guri, Mordechai
Pubblicazione: (2025)
An AI-Driven Framework for Energy-Efficient Environmental Monitoring in Smart Cities Using Edge Intelligence
di: Liu, Yichen, et al.
Pubblicazione: (2026)
di: Liu, Yichen, et al.
Pubblicazione: (2026)
Understanding the Performance and Power of LLM Inferencing on Edge Accelerators
di: Arya, Mayank, et al.
Pubblicazione: (2025)
di: Arya, Mayank, et al.
Pubblicazione: (2025)
Raft Distributed System for Multi-access Edge Computing Sharing Resources
di: Khaliq, Zain, et al.
Pubblicazione: (2024)
di: Khaliq, Zain, et al.
Pubblicazione: (2024)
EdgeProfiler: A Fast Profiling Framework for Lightweight LLMs on Edge Using Analytical Model
di: Pinnock, Alyssa, et al.
Pubblicazione: (2025)
di: Pinnock, Alyssa, et al.
Pubblicazione: (2025)
Keep Your Friends Close: Leveraging Affinity Groups to Accelerate AI Inference Workflows
di: Garrett, Thiago, et al.
Pubblicazione: (2023)
di: Garrett, Thiago, et al.
Pubblicazione: (2023)
GenAI at the Edge: Comprehensive Survey on Empowering Edge Devices
di: Navardi, Mozhgan, et al.
Pubblicazione: (2025)
di: Navardi, Mozhgan, et al.
Pubblicazione: (2025)
AgileLog: A Forkable Shared Log for Agents on Data Streams
di: Bhat, Shreesha G., et al.
Pubblicazione: (2026)
di: Bhat, Shreesha G., et al.
Pubblicazione: (2026)
Many Hands Make Light Work: Accelerating Edge Inference via Multi-Client Collaborative Caching
di: Liang, Wenyi, et al.
Pubblicazione: (2024)
di: Liang, Wenyi, et al.
Pubblicazione: (2024)
Data Sharing at the Edge of the Network: A Disturbance Resilient Multi-modal ITS
di: Mikolasek, Igor, et al.
Pubblicazione: (2024)
di: Mikolasek, Igor, et al.
Pubblicazione: (2024)
Performance Characterization of Containerized DNN Training and Inference on Edge Accelerators
di: K., Prashanthi S., et al.
Pubblicazione: (2023)
di: K., Prashanthi S., et al.
Pubblicazione: (2023)
Fulcrum: Optimizing Concurrent DNN Training and Inferencing on Edge Accelerators
di: K., Prashanthi S., et al.
Pubblicazione: (2025)
di: K., Prashanthi S., et al.
Pubblicazione: (2025)
Mobile Edge Computing
di: Ahmed, Sohaib, et al.
Pubblicazione: (2024)
di: Ahmed, Sohaib, et al.
Pubblicazione: (2024)
Input-Based Ensemble-Learning Method for Dynamic Memory Configuration of Serverless Computing Functions
di: Agarwal, Siddharth, et al.
Pubblicazione: (2024)
di: Agarwal, Siddharth, et al.
Pubblicazione: (2024)
Towards Scalable GPU-Accelerated SNN Training via Temporal Fusion
di: Li, Yanchen, et al.
Pubblicazione: (2024)
di: Li, Yanchen, et al.
Pubblicazione: (2024)
T-MAC: CPU Renaissance via Table Lookup for Low-Bit LLM Deployment on Edge
di: Wei, Jianyu, et al.
Pubblicazione: (2024)
di: Wei, Jianyu, et al.
Pubblicazione: (2024)
Accelerating Latency-Critical Applications with AI-Powered Semi-Automatic Fine-Grained Parallelization on SMT Processors
di: Los, Denis, et al.
Pubblicazione: (2025)
di: Los, Denis, et al.
Pubblicazione: (2025)
Uncertainty Estimation in Multi-Agent Distributed Learning for AI-Enabled Edge Devices
di: Radchenko, Gleb, et al.
Pubblicazione: (2024)
di: Radchenko, Gleb, et al.
Pubblicazione: (2024)
Scheduling Deep Learning Jobs in Multi-Tenant GPU Clusters via Wise Resource Sharing
di: Luo, Yizhou, et al.
Pubblicazione: (2024)
di: Luo, Yizhou, et al.
Pubblicazione: (2024)
HDEE: Heterogeneous Domain Expert Ensemble
di: Ersoy, Oğuzhan, et al.
Pubblicazione: (2025)
di: Ersoy, Oğuzhan, et al.
Pubblicazione: (2025)
SuperSFL: Resource-Heterogeneous Federated Split Learning with Weight-Sharing Super-Networks
di: Asif, Abdullah Al, et al.
Pubblicazione: (2026)
di: Asif, Abdullah Al, et al.
Pubblicazione: (2026)
Conditional Prior-based Non-stationary Channel Estimation Using Accelerated Diffusion Models
di: Mohsin, Muhammad Ahmed, et al.
Pubblicazione: (2025)
di: Mohsin, Muhammad Ahmed, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Minimizing CGYRO HPC Communication Costs in Ensembles with XGYRO by Sharing the Collisional Constant Tensor Structure
di: Sfiligoi, Igor, et al.
Pubblicazione: (2025) -
Accelerated Digital Twin Learning for Edge AI: A Comparison of FPGA and Mobile GPU
di: Xu, Bin, et al.
Pubblicazione: (2025) -
Synergy: Towards On-Body AI via Tiny AI Accelerator Collaboration on Wearables
di: Gong, Taesik, et al.
Pubblicazione: (2023) -
On the Impact of White-box Deployment Strategies for Edge AI on Latency and Model Performance
di: Singh, Jaskirat, et al.
Pubblicazione: (2024) -
Ensemble Method for System Failure Detection Using Large-Scale Telemetry Data
di: Mudgal, Priyanka, et al.
Pubblicazione: (2024)