A QoE-Aware Split Inference Accelerating Algorithm for NOMA-based Edge Intelligence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Xin, Li, Ning, Chen, Quan, Xu, Wenchao, Zhang, Zhaoxin, Guo, Song |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient Edge LLMs Deployment via HessianAware Quantization and CPU GPU Collaborative
von: Zhang, Tuo, et al.
Veröffentlicht: (2025)
von: Zhang, Tuo, et al.
Veröffentlicht: (2025)
Causal-Aware Intelligent QoE Optimization for VR Interaction with Adaptive Keyframe Extraction
von: Zhang, Ziru, et al.
Veröffentlicht: (2025)
von: Zhang, Ziru, et al.
Veröffentlicht: (2025)
QoS-QoE Translation with Large Language Model
von: Yu, Yingjie, et al.
Veröffentlicht: (2026)
von: Yu, Yingjie, et al.
Veröffentlicht: (2026)
QECO: A QoE-Oriented Computation Offloading Algorithm based on Deep Reinforcement Learning for Mobile Edge Computing
von: Rahmaty, Iman, et al.
Veröffentlicht: (2023)
von: Rahmaty, Iman, et al.
Veröffentlicht: (2023)
Generative QoE Modeling: A Lightweight Approach for Telecom Networks
von: Nayar, Vinti, et al.
Veröffentlicht: (2025)
von: Nayar, Vinti, et al.
Veröffentlicht: (2025)
TSKAN: Interpretable Machine Learning for QoE modeling over Time Series Data
von: Singh, Kamal, et al.
Veröffentlicht: (2025)
von: Singh, Kamal, et al.
Veröffentlicht: (2025)
Quantum-based QoE Optimization in Advanced Cellular Networks: Integration and Cloud Gaming Use Case
von: Chaouech, Fatma, et al.
Veröffentlicht: (2025)
von: Chaouech, Fatma, et al.
Veröffentlicht: (2025)
COBRA: Algorithm-Architecture Co-optimized Binary Transformer Accelerator for Edge Inference
von: Qiao, Ye, et al.
Veröffentlicht: (2025)
von: Qiao, Ye, et al.
Veröffentlicht: (2025)
Dora: QoE-Aware Hybrid Parallelism for Distributed Edge AI
von: Jin, Jianli, et al.
Veröffentlicht: (2025)
von: Jin, Jianli, et al.
Veröffentlicht: (2025)
Video QoE Metrics from Encrypted Traffic: Application-agnostic Methodology
von: Berger, Tamir, et al.
Veröffentlicht: (2025)
von: Berger, Tamir, et al.
Veröffentlicht: (2025)
qAttCNN - Self Attention Mechanism for Video QoE Prediction in Encrypted Traffic
von: Sidorov, Michael, et al.
Veröffentlicht: (2026)
von: Sidorov, Michael, et al.
Veröffentlicht: (2026)
CHIME: Chiplet-based Heterogeneous Near-Memory Acceleration for Edge Multimodal LLM Inference
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
von: Chen, Yanru, et al.
Veröffentlicht: (2025)
A Survey on Collaborative DNN Inference for Edge Intelligence
von: Ren, Weiqing, et al.
Veröffentlicht: (2022)
von: Ren, Weiqing, et al.
Veröffentlicht: (2022)
DK-Root: A Joint Data-and-Knowledge-Driven Framework for Root Cause Analysis of QoE Degradations in Mobile Networks
von: Li, Qizhe, et al.
Veröffentlicht: (2025)
von: Li, Qizhe, et al.
Veröffentlicht: (2025)
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism
von: Yan, Jiaming, et al.
Veröffentlicht: (2025)
von: Yan, Jiaming, et al.
Veröffentlicht: (2025)
QoE-Driven Multi-Task Offloading for Semantic-Aware Edge Computing Systems
von: Chen, Xuyang, et al.
Veröffentlicht: (2024)
von: Chen, Xuyang, et al.
Veröffentlicht: (2024)
CoMoE: Collaborative Optimization of Expert Aggregation and Offloading for MoE-based LLMs at Edge
von: Li, Muqing, et al.
Veröffentlicht: (2025)
von: Li, Muqing, et al.
Veröffentlicht: (2025)
A Survey on Trustworthy Edge Intelligence: From Security and Reliability To Transparency and Sustainability
von: Wang, Xiaojie, et al.
Veröffentlicht: (2023)
von: Wang, Xiaojie, et al.
Veröffentlicht: (2023)
CHESTNUT: A QoS Dataset for Mobile Edge Environments
von: Zou, Guobing, et al.
Veröffentlicht: (2024)
von: Zou, Guobing, et al.
Veröffentlicht: (2024)
Towards Integrated Fine-tuning and Inference when Generative AI meets Edge Intelligence
von: Chen, Ning, et al.
Veröffentlicht: (2024)
von: Chen, Ning, et al.
Veröffentlicht: (2024)
Underwater Acoustic Target Recognition based on Smoothness-inducing Regularization and Spectrogram-based Data Augmentation
von: Xu, Ji, et al.
Veröffentlicht: (2023)
von: Xu, Ji, et al.
Veröffentlicht: (2023)
Communication-Efficient Multi-Modal Edge Inference via Uncertainty-Aware Distributed Learning
von: Zhao, Hang, et al.
Veröffentlicht: (2026)
von: Zhao, Hang, et al.
Veröffentlicht: (2026)
Transferable Deployment of Semantic Edge Inference Systems via Unsupervised Domain Adaption
von: Jiao, Weiqiang, et al.
Veröffentlicht: (2025)
von: Jiao, Weiqiang, et al.
Veröffentlicht: (2025)
EdgeFlex-Transformer: Transformer Inference for Edge Devices
von: Mohammad, Shoaib, et al.
Veröffentlicht: (2025)
von: Mohammad, Shoaib, et al.
Veröffentlicht: (2025)
QoS-Nets: Adaptive Approximate Neural Network Inference
von: Trommer, Elias, et al.
Veröffentlicht: (2024)
von: Trommer, Elias, et al.
Veröffentlicht: (2024)
Accelerated Cloud for Artificial Intelligence (ACAI)
von: Chen, Dachi, et al.
Veröffentlicht: (2024)
von: Chen, Dachi, et al.
Veröffentlicht: (2024)
Contract-Driven QoE Auditing for Speech and Singing Services: From MOS Regression to Service Graphs
von: Du, Wenzhang
Veröffentlicht: (2025)
von: Du, Wenzhang
Veröffentlicht: (2025)
SparseDVFS: Sparse-Aware DVFS for Energy-Efficient Edge Inference
von: Zhang, Ziyang, et al.
Veröffentlicht: (2026)
von: Zhang, Ziyang, et al.
Veröffentlicht: (2026)
CAS-Spec: Cascade Adaptive Self-Speculative Decoding for On-the-Fly Lossless Inference Acceleration of LLMs
von: Ning, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Ning, Zhiyuan, et al.
Veröffentlicht: (2025)
ES-GNN: Generalizing Graph Neural Networks Beyond Homophily with Edge Splitting
von: Guo, Jingwei, et al.
Veröffentlicht: (2022)
von: Guo, Jingwei, et al.
Veröffentlicht: (2022)
Harli: SLO-Aware Co-location of LLM Inference and PEFT-based Finetuning on Model-as-a-Service Platforms
von: Xu, Ao, et al.
Veröffentlicht: (2025)
von: Xu, Ao, et al.
Veröffentlicht: (2025)
QoE-Aware and Secure UAV-Aided Rate-Splitting Multiple Access Based Communications
von: Adam, Abuzar B. M., et al.
Veröffentlicht: (2024)
von: Adam, Abuzar B. M., et al.
Veröffentlicht: (2024)
DynaSplit: A Hardware-Software Co-Design Framework for Energy-Aware Inference on Edge
von: May, Daniel, et al.
Veröffentlicht: (2024)
von: May, Daniel, et al.
Veröffentlicht: (2024)
Adaptive Layer Splitting for Wireless LLM Inference in Edge Computing: A Model-Based Reinforcement Learning Approach
von: Chen, Yuxuan, et al.
Veröffentlicht: (2024)
von: Chen, Yuxuan, et al.
Veröffentlicht: (2024)
QoE-based Semantic-Aware Resource Allocation for Multi-Task Networks
von: Yan, Lei, et al.
Veröffentlicht: (2023)
von: Yan, Lei, et al.
Veröffentlicht: (2023)
QoNext: Towards Next-generation QoE for Foundation Models
von: Guo, Yijin, et al.
Veröffentlicht: (2025)
von: Guo, Yijin, et al.
Veröffentlicht: (2025)
Online Resource Allocation for Edge Intelligence with Colocated Model Retraining and Inference
von: Cai, Huaiguang, et al.
Veröffentlicht: (2024)
von: Cai, Huaiguang, et al.
Veröffentlicht: (2024)
Modality-Aware Zero-Shot Pruning and Sparse Attention for Efficient Multimodal Edge Inference
von: Sui, Yueyuan, et al.
Veröffentlicht: (2026)
von: Sui, Yueyuan, et al.
Veröffentlicht: (2026)
Compiler-Assisted Speculative Sampling for Accelerated LLM Inference on Heterogeneous Edge Devices
von: Mesa, Alejandro Ruiz y, et al.
Veröffentlicht: (2026)
von: Mesa, Alejandro Ruiz y, et al.
Veröffentlicht: (2026)
Bayes-Split-Edge: Bayesian Optimization for Constrained Collaborative Inference in Wireless Edge Systems
von: Safaeipour, Fatemeh Zahra, et al.
Veröffentlicht: (2025)
von: Safaeipour, Fatemeh Zahra, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Efficient Edge LLMs Deployment via HessianAware Quantization and CPU GPU Collaborative
von: Zhang, Tuo, et al.
Veröffentlicht: (2025) -
Causal-Aware Intelligent QoE Optimization for VR Interaction with Adaptive Keyframe Extraction
von: Zhang, Ziru, et al.
Veröffentlicht: (2025) -
QoS-QoE Translation with Large Language Model
von: Yu, Yingjie, et al.
Veröffentlicht: (2026) -
QECO: A QoE-Oriented Computation Offloading Algorithm based on Deep Reinforcement Learning for Mobile Edge Computing
von: Rahmaty, Iman, et al.
Veröffentlicht: (2023) -
Generative QoE Modeling: A Lightweight Approach for Telecom Networks
von: Nayar, Vinti, et al.
Veröffentlicht: (2025)