TrimCaching: Parameter-sharing Edge Caching for AI Model Downloading
Fuente:
arXiv
Saved in:
| Main Authors: | Qu, Guanqiao, Lin, Zheng, Chen, Qian, Li, Jian, Liu, Fangming, Chen, Xianhao, Huang, Kaibin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TrimCaching: Parameter-sharing AI Model Caching in Wireless Edge Networks
by: Qu, Guanqiao, et al.
Published: (2024)
by: Qu, Guanqiao, et al.
Published: (2024)
PartialLoading: User Scheduling and Bandwidth Allocation for Parameter-sharing Edge Inference
by: Qu, Guanqiao, et al.
Published: (2025)
by: Qu, Guanqiao, et al.
Published: (2025)
SLIDE: Simultaneous Model Downloading and Inference at the Wireless Network Edge
by: Qu, Guanqiao, et al.
Published: (2025)
by: Qu, Guanqiao, et al.
Published: (2025)
SlimCaching: Edge Caching of Mixture-of-Experts for Distributed Inference
by: Chen, Qian, et al.
Published: (2025)
by: Chen, Qian, et al.
Published: (2025)
Split Learning in 6G Edge Networks
by: Lin, Zheng, et al.
Published: (2023)
by: Lin, Zheng, et al.
Published: (2023)
Mobile Edge Intelligence for Large Language Models: A Contemporary Survey
by: Qu, Guanqiao, et al.
Published: (2024)
by: Qu, Guanqiao, et al.
Published: (2024)
Space-ground Fluid AI for 6G Edge Intelligence
by: Chen, Qian, et al.
Published: (2024)
by: Chen, Qian, et al.
Published: (2024)
SpaceMoE: Towards Orbital General Intelligence with Distributed Mixture-of-Experts Inference
by: Chen, Qian, et al.
Published: (2026)
by: Chen, Qian, et al.
Published: (2026)
SiftMoE: Similarity-Aware Energy-Efficient Expert Selection for Wireless Distributed MoE Inference
by: Chen, Qian, et al.
Published: (2026)
by: Chen, Qian, et al.
Published: (2026)
Parametric-Sensitivity Aware Retransmission for Efficient AI Downloading
by: Zhou, You, et al.
Published: (2026)
by: Zhou, You, et al.
Published: (2026)
FedMeld: A Model-dispersal Federated Learning Framework for Space-ground Integrated Networks
by: Chen, Qian, et al.
Published: (2024)
by: Chen, Qian, et al.
Published: (2024)
Fine-Grained AI Model Caching and Downloading With Coordinated Multipoint Broadcasting in Multi-Cell Edge Networks
by: Fu, Yang, et al.
Published: (2025)
by: Fu, Yang, et al.
Published: (2025)
Pipelining Split Learning in Multi-hop Edge Networks
by: Wei, Wei, et al.
Published: (2025)
by: Wei, Wei, et al.
Published: (2025)
Adaptive Contextual Caching for Mobile Edge Large Language Model Service
by: Liu, Guangyuan, et al.
Published: (2025)
by: Liu, Guangyuan, et al.
Published: (2025)
Joint Model Caching and Resource Allocation in Generative AI-Enabled Wireless Edge Networks
by: Liu, Zhang, et al.
Published: (2024)
by: Liu, Zhang, et al.
Published: (2024)
Cooperative Edge Caching with Large Language Model in Wireless Networks
by: Yang, Ning, et al.
Published: (2026)
by: Yang, Ning, et al.
Published: (2026)
Digital Twin-Enabled Mobility-Aware Cooperative Caching in Vehicular Edge Computing
by: Zeng, Jiahao, et al.
Published: (2026)
by: Zeng, Jiahao, et al.
Published: (2026)
Coffee: Cost-Effective Edge Caching for 360 Degree Live Video Streaming
by: Li, Chen, et al.
Published: (2023)
by: Li, Chen, et al.
Published: (2023)
Joint Optimization of DNN Model Caching and Request Routing in Mobile Edge Computing
by: Qiu, Shuting, et al.
Published: (2025)
by: Qiu, Shuting, et al.
Published: (2025)
FluxShard: Motion-Aware Feature Cache Reuse for Collaborative Video Analytics in Mobile Edge Computing
by: Guan, Xiuxian, et al.
Published: (2026)
by: Guan, Xiuxian, et al.
Published: (2026)
Semantic-Aware Caching for Efficient Image Generation in Edge Computing
by: Cui, Hanshuai, et al.
Published: (2025)
by: Cui, Hanshuai, et al.
Published: (2025)
Joint Cache Placement and Routing in Satellite-Terrestrial Edge Computing Network: A GNN-Enabled DRL Approach
by: Zheng, Yuhao, et al.
Published: (2025)
by: Zheng, Yuhao, et al.
Published: (2025)
CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
by: Liu, Yuhan, et al.
Published: (2023)
by: Liu, Yuhan, et al.
Published: (2023)
VNF-Cache: An In-Network Key-Value Store Cache Based on Network Function Virtualization
by: Farias, Bruno E., et al.
Published: (2025)
by: Farias, Bruno E., et al.
Published: (2025)
Personalized Federated Deep Reinforcement Learning for Heterogeneous Edge Content Caching Networks
by: Li, Zhen, et al.
Published: (2024)
by: Li, Zhen, et al.
Published: (2024)
Digital Twin-Assisted Data-Driven Optimization for Reliable Edge Caching in Wireless Networks
by: Zhang, Zifan, et al.
Published: (2024)
by: Zhang, Zifan, et al.
Published: (2024)
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading
by: Xu, Minrui, et al.
Published: (2025)
by: Xu, Minrui, et al.
Published: (2025)
Learning Cache Coherence Traffic for NoC Routing Design
by: Xiong, Guochu, et al.
Published: (2025)
by: Xiong, Guochu, et al.
Published: (2025)
Fundamentals of Caching Layered Data objects
by: Bari, Agrim, et al.
Published: (2025)
by: Bari, Agrim, et al.
Published: (2025)
MemorAI: Energy-Efficient Last-Level Cache Memory Optimization for Virtualized RANs
by: Hidalgo, Ethan Sanchez, et al.
Published: (2024)
by: Hidalgo, Ethan Sanchez, et al.
Published: (2024)
LLM-Empowered Cooperative Content Caching in Vehicular Fog Caching-Assisted Platoon Networks
by: Tan, Bowen, et al.
Published: (2026)
by: Tan, Bowen, et al.
Published: (2026)
cRVR: A Stackelberg Game Approach for Joint Privacy-Aware Video Requesting and Edge Caching
by: Zhang, Xianzhi, et al.
Published: (2023)
by: Zhang, Xianzhi, et al.
Published: (2023)
RRTO: A High-Performance Transparent Offloading System for Model Inference in Mobile Edge Computing
by: Sun, Zekai, et al.
Published: (2025)
by: Sun, Zekai, et al.
Published: (2025)
VEC-Sim: A Simulation Platform for Evaluating Service Caching and Computation Offloading Policies in Vehicular Edge Networks
by: Wu, Fan, et al.
Published: (2024)
by: Wu, Fan, et al.
Published: (2024)
Content Caching Methods in Named Data Networks
by: Chaudhary, Pankaj, et al.
Published: (2026)
by: Chaudhary, Pankaj, et al.
Published: (2026)
Bandwidth Efficient Cache Selection and Content Advertisement
by: Cohen, Itamar
Published: (2024)
by: Cohen, Itamar
Published: (2024)
Federated Learning Assisted Edge Caching Scheme Based on Lightweight Architecture DDPM
by: Li, Xun, et al.
Published: (2025)
by: Li, Xun, et al.
Published: (2025)
Modeling and Optimizing Latency for Delayed Hit Caching with Stochastic Miss Latency
by: Jiang, Bowen, et al.
Published: (2025)
by: Jiang, Bowen, et al.
Published: (2025)
NetMCP: Network-Aware Model Context Protocol Platform for LLM Capability Extension
by: Li, Enhan, et al.
Published: (2025)
by: Li, Enhan, et al.
Published: (2025)
CacheMamba: Popularity Prediction for Mobile Edge Caching Networks via Selective State Spaces
by: Kianfar, Ghazaleh, et al.
Published: (2025)
by: Kianfar, Ghazaleh, et al.
Published: (2025)
Similar Items
-
TrimCaching: Parameter-sharing AI Model Caching in Wireless Edge Networks
by: Qu, Guanqiao, et al.
Published: (2024) -
PartialLoading: User Scheduling and Bandwidth Allocation for Parameter-sharing Edge Inference
by: Qu, Guanqiao, et al.
Published: (2025) -
SLIDE: Simultaneous Model Downloading and Inference at the Wireless Network Edge
by: Qu, Guanqiao, et al.
Published: (2025) -
SlimCaching: Edge Caching of Mixture-of-Experts for Distributed Inference
by: Chen, Qian, et al.
Published: (2025) -
Split Learning in 6G Edge Networks
by: Lin, Zheng, et al.
Published: (2023)