Conformer-Based Speech Recognition On Extreme Edge-Computing Devices
Fuente:
arXiv
Guardado en:
| Autores principales: | Xu, Mingbin, Jin, Alex, Wang, Sicheng, Su, Mu, Ng, Tim, Mason, Henry, Han, Shiyi, Lei, Zhihong, Deng, Yaqiao, Huang, Zhen, Krishnamoorthy, Mahesh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Spatiotemporal Analysis of Parallelized Computing at the Extreme Edge
por: Nabil, Yasser, et al.
Publicado: (2025)
por: Nabil, Yasser, et al.
Publicado: (2025)
Performance Characterization of Containers in Edge Computing
por: Gupta, Ragini, et al.
Publicado: (2025)
por: Gupta, Ragini, et al.
Publicado: (2025)
Spatiotemporal Non-Uniformity-Aware Online Task Scheduling in Collaborative Edge Computing for Industrial Internet of Things
por: Li, Yang, et al.
Publicado: (2025)
por: Li, Yang, et al.
Publicado: (2025)
Automating Energy-Efficient GPU Kernel Generation: A Fast Search-Based Compilation Approach
por: Zhang, Yijia, et al.
Publicado: (2024)
por: Zhang, Yijia, et al.
Publicado: (2024)
Modeling Tradeoffs between mobility, cost, and performance in Edge Computing
por: Waseem, Muhammad Danish, et al.
Publicado: (2026)
por: Waseem, Muhammad Danish, et al.
Publicado: (2026)
Latency and Privacy-Aware Resource Allocation in Vehicular Edge Computing
por: Ahmadvand, Hossein, et al.
Publicado: (2025)
por: Ahmadvand, Hossein, et al.
Publicado: (2025)
FlexQuant: Elastic Quantization Framework for Locally Hosted LLM on Edge Devices
por: Chai, Yuji, et al.
Publicado: (2025)
por: Chai, Yuji, et al.
Publicado: (2025)
CarbonCP: Carbon-Aware DNN Partitioning with Conformal Prediction for Sustainable Edge Intelligence
por: Ke, Hongyu, et al.
Publicado: (2024)
por: Ke, Hongyu, et al.
Publicado: (2024)
A Structure-Aware Framework for Learning Device Placements on Computation Graphs
por: Duan, Shukai, et al.
Publicado: (2024)
por: Duan, Shukai, et al.
Publicado: (2024)
Characterizing and Optimizing Realistic Workloads on a Commercial Compute-in-SRAM Device
por: Zhang, Niansong, et al.
Publicado: (2025)
por: Zhang, Niansong, et al.
Publicado: (2025)
Resource-Efficient RGB-Only Action Recognition for Edge Deployment
por: Yoon, Dongsik, et al.
Publicado: (2026)
por: Yoon, Dongsik, et al.
Publicado: (2026)
HPC Application Parameter Autotuning on Edge Devices: A Bandit Learning Approach
por: Hossain, Abrar, et al.
Publicado: (2025)
por: Hossain, Abrar, et al.
Publicado: (2025)
CarbonCall: Sustainability-Aware Function Calling for Large Language Models on Edge Devices
por: Paramanayakam, Varatheepan, et al.
Publicado: (2025)
por: Paramanayakam, Varatheepan, et al.
Publicado: (2025)
Memory Analysis on the Training Course of DeepSeek Models
por: Zhang, Ping, et al.
Publicado: (2025)
por: Zhang, Ping, et al.
Publicado: (2025)
Enhancing CTC-based speech recognition with diverse modeling units
por: Han, Shiyi, et al.
Publicado: (2024)
por: Han, Shiyi, et al.
Publicado: (2024)
Application Research On Real-Time Perception Of Device Performance Status
por: Wang, Zhe, et al.
Publicado: (2024)
por: Wang, Zhe, et al.
Publicado: (2024)
Cloud Computing Energy Consumption Prediction Based on Kernel Extreme Learning Machine Algorithm Improved by Vector Weighted Average Algorithm
por: Wang, Yuqing, et al.
Publicado: (2025)
por: Wang, Yuqing, et al.
Publicado: (2025)
Modeling Interfering Sources in Shared Queues for Timely Computations in Edge Computing Systems
por: Akar, Nail, et al.
Publicado: (2024)
por: Akar, Nail, et al.
Publicado: (2024)
CoFormer: Collaborating with Heterogeneous Edge Devices for Scalable Transformer Inference
por: Xu, Guanyu, et al.
Publicado: (2025)
por: Xu, Guanyu, et al.
Publicado: (2025)
Less is More: Optimizing Function Calling for LLM Execution on Edge Devices
por: Paramanayakam, Varatheepan, et al.
Publicado: (2024)
por: Paramanayakam, Varatheepan, et al.
Publicado: (2024)
RooflineBench: A Benchmarking Framework for On-Device LLMs via Roofline Analysis
por: Bi, Zhen, et al.
Publicado: (2026)
por: Bi, Zhen, et al.
Publicado: (2026)
ISO: Overlap of Computation and Communication within Seqenence For LLM Inference
por: Xiao, Bin, et al.
Publicado: (2024)
por: Xiao, Bin, et al.
Publicado: (2024)
Characterize LSM-tree Compaction Performance via On-Device LLM Inference
por: Ding, Jiabiao, et al.
Publicado: (2026)
por: Ding, Jiabiao, et al.
Publicado: (2026)
Collaborative Processing for Multi-Tenant Inference on Memory-Constrained Edge TPUs
por: Ng, Nathan, et al.
Publicado: (2026)
por: Ng, Nathan, et al.
Publicado: (2026)
AFarePart: Accuracy-aware Fault-resilient Partitioner for DNN Edge Accelerators
por: Debnath, Mukta, et al.
Publicado: (2025)
por: Debnath, Mukta, et al.
Publicado: (2025)
Ecoscape: Fault Tolerance Benchmark for Adaptive Remediation Strategies in Real-Time Edge ML
por: Reiter, Hendrik, et al.
Publicado: (2025)
por: Reiter, Hendrik, et al.
Publicado: (2025)
DEER: Deep Runahead for Instruction Prefetching on Modern Mobile Workloads
por: Vahdatniya, Parmida, et al.
Publicado: (2025)
por: Vahdatniya, Parmida, et al.
Publicado: (2025)
Unikernels vs. Containers: A Runtime-Level Performance Comparison for Resource-Constrained Edge Workloads
por: Dinh-Tuan, Hai
Publicado: (2025)
por: Dinh-Tuan, Hai
Publicado: (2025)
Two-Timescale Dynamic Service Deployment and Task Scheduling with Spatiotemporal Collaboration in Mobile Edge Networks
por: Li, Yang, et al.
Publicado: (2025)
por: Li, Yang, et al.
Publicado: (2025)
XRFlux: Virtual Reality Benchmark for Edge Caching Systems
por: Alfares, Nader, et al.
Publicado: (2024)
por: Alfares, Nader, et al.
Publicado: (2024)
Redundant Array Computation Elimination
por: Wang, Zixuan, et al.
Publicado: (2025)
por: Wang, Zixuan, et al.
Publicado: (2025)
Performance of Confidential Computing GPUs
por: Ibarra, Antonio Martínez, et al.
Publicado: (2025)
por: Ibarra, Antonio Martínez, et al.
Publicado: (2025)
Rethinking Temporal Models for TinyML: LSTM versus 1D-CNN in Resource-Constrained Devices
por: Saha, Bidyut, et al.
Publicado: (2026)
por: Saha, Bidyut, et al.
Publicado: (2026)
Achieving Consistent and Comparable CPU Evaluation
por: Wang, Chenxi, et al.
Publicado: (2024)
por: Wang, Chenxi, et al.
Publicado: (2024)
Attributing the System's Overall Effect to its Components
por: Wang, Chenxi, et al.
Publicado: (2026)
por: Wang, Chenxi, et al.
Publicado: (2026)
MoE-Infinity: Efficient MoE Inference on Personal Machines with Sparsity-Aware Expert Cache
por: Xue, Leyang, et al.
Publicado: (2024)
por: Xue, Leyang, et al.
Publicado: (2024)
Multi-Dimensional Autoscaling of Stream Processing Services on Edge Devices
por: Sedlak, Boris, et al.
Publicado: (2025)
por: Sedlak, Boris, et al.
Publicado: (2025)
CXL-Interference: Analysis and Characterization in Modern Computer Systems
por: Mao, Shunyu, et al.
Publicado: (2024)
por: Mao, Shunyu, et al.
Publicado: (2024)
Resource Management Schemes for Cloud-Native Platforms with Computing Containers of Docker and Kubernetes
por: Mao, Ying, et al.
Publicado: (2020)
por: Mao, Ying, et al.
Publicado: (2020)
Computational Complexity-Constrained Spectral Efficiency Analysis for 6G Waveforms
por: Queiroz, Saulo, et al.
Publicado: (2024)
por: Queiroz, Saulo, et al.
Publicado: (2024)
Ejemplares similares
-
Spatiotemporal Analysis of Parallelized Computing at the Extreme Edge
por: Nabil, Yasser, et al.
Publicado: (2025) -
Performance Characterization of Containers in Edge Computing
por: Gupta, Ragini, et al.
Publicado: (2025) -
Spatiotemporal Non-Uniformity-Aware Online Task Scheduling in Collaborative Edge Computing for Industrial Internet of Things
por: Li, Yang, et al.
Publicado: (2025) -
Automating Energy-Efficient GPU Kernel Generation: A Fast Search-Based Compilation Approach
por: Zhang, Yijia, et al.
Publicado: (2024) -
Modeling Tradeoffs between mobility, cost, and performance in Edge Computing
por: Waseem, Muhammad Danish, et al.
Publicado: (2026)