Synergy: Towards On-Body AI via Tiny AI Accelerator Collaboration on Wearables
Fuente:
arXiv
Guardado en:
| Autores principales: | Gong, Taesik, Jang, Si Young, Acer, Utku Günay, Kawsar, Fahim, Min, Chulhong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
An AI-Native Runtime for Multi-Wearable Environments
por: Min, Chulhong, et al.
Publicado: (2024)
por: Min, Chulhong, et al.
Publicado: (2024)
Enabling Cross-Camera Collaboration for Video Analytics on Distributed Smart Cameras
por: Min, Chulhong, et al.
Publicado: (2024)
por: Min, Chulhong, et al.
Publicado: (2024)
Salted Inference: Enhancing Privacy while Maintaining Efficiency of Split Inference in Mobile Computing
por: Malekzadeh, Mohammad, et al.
Publicado: (2023)
por: Malekzadeh, Mohammad, et al.
Publicado: (2023)
Smaller, Smarter, Closer: The Edge of Collaborative Generative AI
por: Morabito, Roberto, et al.
Publicado: (2025)
por: Morabito, Roberto, et al.
Publicado: (2025)
DEX: Data Channel Extension for Efficient CNN Inference on Tiny AI Accelerators
por: Gong, Taesik, et al.
Publicado: (2024)
por: Gong, Taesik, et al.
Publicado: (2024)
SynergAI: Edge-to-Cloud Synergy for Architecture-Driven High-Performance Orchestration for AI Inference
por: Stathopoulou, Foteini, et al.
Publicado: (2025)
por: Stathopoulou, Foteini, et al.
Publicado: (2025)
EaCO: Resource Sharing Dynamics and Its Impact on Energy Efficiency for DNN Training
por: Haghshenas, Kawsar, et al.
Publicado: (2024)
por: Haghshenas, Kawsar, et al.
Publicado: (2024)
ML-ECS: A Collaborative Multimodal Learning Framework for Edge-Cloud Synergies
por: Liu, Yuze, et al.
Publicado: (2026)
por: Liu, Yuze, et al.
Publicado: (2026)
Accelerating End-Cloud Collaborative Inference via Near Bubble-free Pipeline Optimization
por: Gao, Luyao, et al.
Publicado: (2024)
por: Gao, Luyao, et al.
Publicado: (2024)
Tiny Deep Ensemble: Uncertainty Estimation in Edge AI Accelerators via Ensembling Normalization Layers with Shared Weights
por: Ahmed, Soyed Tuhin, et al.
Publicado: (2024)
por: Ahmed, Soyed Tuhin, et al.
Publicado: (2024)
Gaia: Hybrid Hardware Acceleration for Serverless AI in the 3D Compute Continuum
por: Reisecker, Maximilian, et al.
Publicado: (2025)
por: Reisecker, Maximilian, et al.
Publicado: (2025)
Collaborative Inference Acceleration with Non-Penetrative Tensor Partitioning
por: Liu, Zhibang, et al.
Publicado: (2025)
por: Liu, Zhibang, et al.
Publicado: (2025)
exa-AMD: A Scalable Workflow for Accelerating AI-Assisted Materials Discovery and Design
por: Moraru, Maxim, et al.
Publicado: (2025)
por: Moraru, Maxim, et al.
Publicado: (2025)
Many Hands Make Light Work: Accelerating Edge Inference via Multi-Client Collaborative Caching
por: Liang, Wenyi, et al.
Publicado: (2024)
por: Liang, Wenyi, et al.
Publicado: (2024)
Efficient Local-to-Global Collaborative Perception via Joint Communication and Computation Optimization
por: Zhang, Hui, et al.
Publicado: (2026)
por: Zhang, Hui, et al.
Publicado: (2026)
BandPilot: Towards Performance- and Contention-Aware GPU Dispatching in AI Clusters
por: Zhang, Kunming, et al.
Publicado: (2025)
por: Zhang, Kunming, et al.
Publicado: (2025)
LIME:Accelerating Collaborative Lossless LLM Inference on Memory-Constrained Edge Devices
por: Sun, Mingyu, et al.
Publicado: (2025)
por: Sun, Mingyu, et al.
Publicado: (2025)
KAITIAN: A Unified Communication Framework for Enabling Efficient Collaboration Across Heterogeneous Accelerators in Embodied AI Systems
por: Lin, Jieke, et al.
Publicado: (2025)
por: Lin, Jieke, et al.
Publicado: (2025)
Parallel Collaborative ADMM Privacy Computing and Adaptive GPU Acceleration for Distributed Edge Networks
por: Xia, Mengchun, et al.
Publicado: (2026)
por: Xia, Mengchun, et al.
Publicado: (2026)
SCARIF: Towards Carbon Modeling of Cloud Servers with Accelerators
por: Ji, Shixin, et al.
Publicado: (2024)
por: Ji, Shixin, et al.
Publicado: (2024)
SGCP: A Self-Organized Game-Theoretic Framework For Collaborative Perception
por: Gong, Zechuan, et al.
Publicado: (2026)
por: Gong, Zechuan, et al.
Publicado: (2026)
Agentic AI Workload Characteristics
por: Yuan, Yichao, et al.
Publicado: (2026)
por: Yuan, Yichao, et al.
Publicado: (2026)
Cloud abstractions for AI workloads
por: Canini, Marco, et al.
Publicado: (2025)
por: Canini, Marco, et al.
Publicado: (2025)
Efficient Unified Caching for Accelerating Heterogeneous AI Workloads
por: Wang, Tianze, et al.
Publicado: (2025)
por: Wang, Tianze, et al.
Publicado: (2025)
FLEX: Leveraging FPGA-CPU Synergy for Mixed-Cell-Height Legalization Acceleration
por: Liu, Xingyu, et al.
Publicado: (2025)
por: Liu, Xingyu, et al.
Publicado: (2025)
Syncopate: Efficient Multi-GPU AI Kernels via Automatic Chunk-Centric Compute-Communication Overlap
por: Qiang, Xinwei, et al.
Publicado: (2026)
por: Qiang, Xinwei, et al.
Publicado: (2026)
AI Surrogate Model for Distributed Computing Workloads
por: Park, David K., et al.
Publicado: (2024)
por: Park, David K., et al.
Publicado: (2024)
BanaServe: Unified KV Cache and Dynamic Module Migration for Balancing Disaggregated LLM Serving in AI Infrastructure
por: He, Yiyuan, et al.
Publicado: (2025)
por: He, Yiyuan, et al.
Publicado: (2025)
Accelerating the Delivery of Data Services over Uncertain Mobile Crowdsensing Networks
por: Liwang, Minghui, et al.
Publicado: (2022)
por: Liwang, Minghui, et al.
Publicado: (2022)
Accelerating Heterogeneous Tensor Parallelism via Flexible Workload Control
por: Wang, Zhigang, et al.
Publicado: (2024)
por: Wang, Zhigang, et al.
Publicado: (2024)
Accelerating OpenPangu Inference on NPU via Speculative Decoding
por: Dai, Yuntao, et al.
Publicado: (2026)
por: Dai, Yuntao, et al.
Publicado: (2026)
Parallel Scan on Ascend AI Accelerators
por: Wróblewski, Bartłomiej, et al.
Publicado: (2025)
por: Wróblewski, Bartłomiej, et al.
Publicado: (2025)
EdgeShard: Efficient LLM Inference via Collaborative Edge Computing
por: Zhang, Mingjin, et al.
Publicado: (2024)
por: Zhang, Mingjin, et al.
Publicado: (2024)
AI-coupled HPC Workflow Applications, Middleware and Performance
por: Brewer, Wes, et al.
Publicado: (2024)
por: Brewer, Wes, et al.
Publicado: (2024)
6G EdgeAI: Performance Evaluation and Analysis
por: Yang, Chien-Sheng, et al.
Publicado: (2025)
por: Yang, Chien-Sheng, et al.
Publicado: (2025)
RHAPSODY: Execution of Hybrid AI-HPC Workflows at Scale
por: Alsaadi, Aymen, et al.
Publicado: (2025)
por: Alsaadi, Aymen, et al.
Publicado: (2025)
6G Infrastructures for Edge AI: An Analytical Perspective
por: Horvath, Kurt, et al.
Publicado: (2025)
por: Horvath, Kurt, et al.
Publicado: (2025)
HyperParallel: A Supernode-Affinity AI Framework
por: Zhang, Xin, et al.
Publicado: (2026)
por: Zhang, Xin, et al.
Publicado: (2026)
Towards Resource-Efficient Compound AI Systems
por: Chaudhry, Gohar Irfan, et al.
Publicado: (2025)
por: Chaudhry, Gohar Irfan, et al.
Publicado: (2025)
OServe: Accelerating LLM Serving via Spatial-Temporal Workload Orchestration
por: Jiang, Youhe, et al.
Publicado: (2026)
por: Jiang, Youhe, et al.
Publicado: (2026)
Ejemplares similares
-
An AI-Native Runtime for Multi-Wearable Environments
por: Min, Chulhong, et al.
Publicado: (2024) -
Enabling Cross-Camera Collaboration for Video Analytics on Distributed Smart Cameras
por: Min, Chulhong, et al.
Publicado: (2024) -
Salted Inference: Enhancing Privacy while Maintaining Efficiency of Split Inference in Mobile Computing
por: Malekzadeh, Mohammad, et al.
Publicado: (2023) -
Smaller, Smarter, Closer: The Edge of Collaborative Generative AI
por: Morabito, Roberto, et al.
Publicado: (2025) -
DEX: Data Channel Extension for Efficient CNN Inference on Tiny AI Accelerators
por: Gong, Taesik, et al.
Publicado: (2024)