Semantic-Aware LLM Orchestration for Proactive Resource Management in Predictive Digital Twin Vehicular Networks
Fuente:
arXiv
Guardado en:
| Autor principal: | Ahmadpanah, Seyed Hossein |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
RouterWise: Joint Resource Allocation and Routing for Latency-Aware Multi-Model LLM Serving
por: Kasnavieh, Hossein Hosseini, et al.
Publicado: (2026)
por: Kasnavieh, Hossein Hosseini, et al.
Publicado: (2026)
Resource Allocation Driven by Large Models in Future Semantic-Aware Networks
por: Zhang, Haijun, et al.
Publicado: (2025)
por: Zhang, Haijun, et al.
Publicado: (2025)
On Effectiveness of Graph Neural Network Architectures for Network Digital Twins (NDTs)
por: Zacarias, Iulisloi, et al.
Publicado: (2025)
por: Zacarias, Iulisloi, et al.
Publicado: (2025)
Distributed Simulation for Digital Twins of Large-Scale Real-World DiffServ-Based Networks
por: Huang, Zhuoyao, et al.
Publicado: (2024)
por: Huang, Zhuoyao, et al.
Publicado: (2024)
Plotinus: A Satellite Internet Digital Twin System
por: Gao, Yue, et al.
Publicado: (2024)
por: Gao, Yue, et al.
Publicado: (2024)
Temporal-Aware GPU Resource Allocation for Distributed LLM Inference via Reinforcement Learning
por: Du, Chengze, et al.
Publicado: (2025)
por: Du, Chengze, et al.
Publicado: (2025)
OrchestrRL: Dynamic Compute and Network Orchestration for Disaggregated RL
por: Tan, Xin, et al.
Publicado: (2026)
por: Tan, Xin, et al.
Publicado: (2026)
Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration
por: Luo, Haoxiang, et al.
Publicado: (2025)
por: Luo, Haoxiang, et al.
Publicado: (2025)
Blockchain-based Trust Management in Security Credential Management System for Vehicular Network
por: Byun, SangHyun, et al.
Publicado: (2025)
por: Byun, SangHyun, et al.
Publicado: (2025)
Network Digital Untwinning: Towards Backward Optimization of Digital Twins
por: Zhang, Zifan, et al.
Publicado: (2026)
por: Zhang, Zifan, et al.
Publicado: (2026)
An Auction-Based Mechanism for Optimal Task Allocation and Resource Aware Containerization
por: kumar, Ramakant
Publicado: (2026)
por: kumar, Ramakant
Publicado: (2026)
A Survey on Resource Management in Joint Communication and Computing-Embedded SAGIN
por: Chen, Qian, et al.
Publicado: (2024)
por: Chen, Qian, et al.
Publicado: (2024)
HALO: Semantic-Aware Distributed LLM Inference in Lossy Edge Network
por: Zheng, Peirong, et al.
Publicado: (2026)
por: Zheng, Peirong, et al.
Publicado: (2026)
Surviving the Edge: Federated Learning under Networking and Resource Constraints
por: Mwanje, Mike, et al.
Publicado: (2026)
por: Mwanje, Mike, et al.
Publicado: (2026)
Recursive Offloading for LLM Serving in Multi-tier Networks
por: Wu, Zhiyuan, et al.
Publicado: (2025)
por: Wu, Zhiyuan, et al.
Publicado: (2025)
Legible Consensus: Topology-Aware Quorum Geometry for Asymmetric Networks
por: Mason, Tony
Publicado: (2026)
por: Mason, Tony
Publicado: (2026)
Contention-Aware Microservice Deployment in Collaborative Mobile Edge Networks
por: Ge, Xinlei, et al.
Publicado: (2024)
por: Ge, Xinlei, et al.
Publicado: (2024)
SANSee: A Physical-layer Semantic-aware Networking Framework for Distributed Wireless Sensing
por: Zhu, Huixiang, et al.
Publicado: (2024)
por: Zhu, Huixiang, et al.
Publicado: (2024)
Securing Distributed Network Digital Twin Systems Against Model Poisoning Attacks
por: Zhang, Zifan, et al.
Publicado: (2024)
por: Zhang, Zifan, et al.
Publicado: (2024)
Dynamic Edge Server Selection in Time-Varying Environments: A Reliability-Aware Predictive Approach
por: Burbano, Jaime Sebastian, et al.
Publicado: (2025)
por: Burbano, Jaime Sebastian, et al.
Publicado: (2025)
Risk-Aware and Stable Edge Server Selection Under Network Latency SLOs
por: Liyanage, Mohan, et al.
Publicado: (2026)
por: Liyanage, Mohan, et al.
Publicado: (2026)
Enhancing Digital Forensics Readiness In Big Data Wireless Medical Networks: A Secure Decentralised Framework
por: Mpungu, Cephas, et al.
Publicado: (2024)
por: Mpungu, Cephas, et al.
Publicado: (2024)
Context-Aware Orchestration of Energy-Efficient Gossip Learning Schemes
por: Dinani, Mina Aghaei, et al.
Publicado: (2024)
por: Dinani, Mina Aghaei, et al.
Publicado: (2024)
FedSkipTwin: Digital-Twin-Guided Client Skipping for Communication-Efficient Federated Learning
por: Commey, Daniel, et al.
Publicado: (2025)
por: Commey, Daniel, et al.
Publicado: (2025)
A Task Decomposition and Planning Framework for Efficient LLM Inference in AI-Enabled WiFi-Offload Networks
por: Han, Mingqi, et al.
Publicado: (2026)
por: Han, Mingqi, et al.
Publicado: (2026)
GORGO: Maximizing KV-Cache Reuse While Minimizing Network Latency in Cross-Region LLM Load Balancing
por: Toniolo, Alessio Ricci, et al.
Publicado: (2026)
por: Toniolo, Alessio Ricci, et al.
Publicado: (2026)
Enabling Intelligent Vehicular Networks Through Distributed Learning in the Non-Terrestrial Networks 6G Vision
por: Naseh, David, et al.
Publicado: (2023)
por: Naseh, David, et al.
Publicado: (2023)
PerLLM: Personalized Inference Scheduling with Edge-Cloud Collaboration for Diverse LLM Services
por: Yang, Zheming, et al.
Publicado: (2024)
por: Yang, Zheming, et al.
Publicado: (2024)
Efficient Data Management for IPFS dApps
por: Estrada-Galiñanes, Vero, et al.
Publicado: (2024)
por: Estrada-Galiñanes, Vero, et al.
Publicado: (2024)
Varuna: Enabling Failure-Type Aware RDMA Failover
por: Wang, Xiaoyang, et al.
Publicado: (2026)
por: Wang, Xiaoyang, et al.
Publicado: (2026)
Multi-stage Flow Scheduling for LLM Serving
por: Sun, Yijun, et al.
Publicado: (2026)
por: Sun, Yijun, et al.
Publicado: (2026)
SARS: A Resource Selection Algorithm for Autonomous Driving Tasks in Heterogeneous Mobile Edge Computing
por: Zakerian, Reza, et al.
Publicado: (2024)
por: Zakerian, Reza, et al.
Publicado: (2024)
A Hybrid Cloud Management Plane for Data Processing Pipelines
por: Babu, Vignesh, et al.
Publicado: (2025)
por: Babu, Vignesh, et al.
Publicado: (2025)
Carbon-Aware Temporal Data Transfer Scheduling Across Cloud Datacenters
por: Rodrigues, Elvis, et al.
Publicado: (2025)
por: Rodrigues, Elvis, et al.
Publicado: (2025)
When Digital Twin Meets 6G: Concepts, Obstacles, and Research Prospects
por: Liu, Wenshuai, et al.
Publicado: (2024)
por: Liu, Wenshuai, et al.
Publicado: (2024)
Future Resource Bank for ISAC: Achieving Fast and Stable Win-Win Matching for Both Individuals and Coalitions
por: Qi, Houyi, et al.
Publicado: (2025)
por: Qi, Houyi, et al.
Publicado: (2025)
QoS-Aware Load Balancing in the Computing Continuum via Multi-Player Bandits
por: Čilić, Ivan, et al.
Publicado: (2025)
por: Čilić, Ivan, et al.
Publicado: (2025)
An Online Fragmentation-Aware GPU Scheduler for Multi-Tenant MIG-based Clouds
por: Zambianco, Marco, et al.
Publicado: (2025)
por: Zambianco, Marco, et al.
Publicado: (2025)
CCL-Bench 1.0: A Trace-Based Benchmark for LLM Infrastructure
por: Ding, Eric, et al.
Publicado: (2026)
por: Ding, Eric, et al.
Publicado: (2026)
RailX: A Flexible, Scalable, and Low-Cost Network Architecture for Hyper-Scale LLM Training Systems
por: Feng, Yinxiao, et al.
Publicado: (2025)
por: Feng, Yinxiao, et al.
Publicado: (2025)
Ejemplares similares
-
RouterWise: Joint Resource Allocation and Routing for Latency-Aware Multi-Model LLM Serving
por: Kasnavieh, Hossein Hosseini, et al.
Publicado: (2026) -
Resource Allocation Driven by Large Models in Future Semantic-Aware Networks
por: Zhang, Haijun, et al.
Publicado: (2025) -
On Effectiveness of Graph Neural Network Architectures for Network Digital Twins (NDTs)
por: Zacarias, Iulisloi, et al.
Publicado: (2025) -
Distributed Simulation for Digital Twins of Large-Scale Real-World DiffServ-Based Networks
por: Huang, Zhuoyao, et al.
Publicado: (2024) -
Plotinus: A Satellite Internet Digital Twin System
por: Gao, Yue, et al.
Publicado: (2024)