Scene-Aware Latency Estimation for Microservices via Multi-Scale Graph Fusion
Fuente:
arXiv
Guardado en:
| Autores principales: | Sun, Zhichao, Zhao, Hailiang, Chow, Kingsum |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reliable Microservice Tail Latency Prediction via Decoupled Dual-Stream Learning and Gradient Modulation
por: Qian, Wenzhuo, et al.
Publicado: (2025)
por: Qian, Wenzhuo, et al.
Publicado: (2025)
HGraphScale: Hierarchical Graph Learning for Autoscaling Microservice Applications in Container-based Cloud Computing
por: Fang, Zhengxin, et al.
Publicado: (2025)
por: Fang, Zhengxin, et al.
Publicado: (2025)
StatuScale: Status-aware and Elastic Scaling Strategy for Microservice Applications
por: Wen, Linfeng, et al.
Publicado: (2024)
por: Wen, Linfeng, et al.
Publicado: (2024)
Resilient Auto-Scaling of Microservice Architectures with Efficient Resource Management
por: Ahmad, Hussain, et al.
Publicado: (2025)
por: Ahmad, Hussain, et al.
Publicado: (2025)
Collaborative Evolution of Intelligent Agents in Large-Scale Microservice Systems
por: Li, Yilin, et al.
Publicado: (2025)
por: Li, Yilin, et al.
Publicado: (2025)
SeaLLM: Service-Aware and Latency-Optimized Resource Sharing for Large Language Model Inference
por: Zhao, Yihao, et al.
Publicado: (2025)
por: Zhao, Yihao, et al.
Publicado: (2025)
Data-Locality-Aware Task Assignment and Scheduling for Distributed Job Executions
por: Zhao, Hailiang, et al.
Publicado: (2024)
por: Zhao, Hailiang, et al.
Publicado: (2024)
Hestia: Hyperthread-Level Scheduling for Cloud Microservices with Interference-Aware Attention
por: Yang, Dingyu, et al.
Publicado: (2026)
por: Yang, Dingyu, et al.
Publicado: (2026)
Modular Foundation Model Inference at the Edge: Network-Aware Microservice Optimization
por: Zhu, Juan, et al.
Publicado: (2026)
por: Zhu, Juan, et al.
Publicado: (2026)
Carbon-Aware Microservice Deployment for Optimal User Experience on a Budget
por: Kreutz, Kevin, et al.
Publicado: (2025)
por: Kreutz, Kevin, et al.
Publicado: (2025)
CascadeInfer: Length-Aware Scheduling of LLM Serving with Low Latency and Load Balancing
por: Yuan, Yitao, et al.
Publicado: (2025)
por: Yuan, Yitao, et al.
Publicado: (2025)
Beyond Microservices: Testing Web-Scale RCA Methods on GPU-Driven LLM Workloads
por: Scheinert, Dominik, et al.
Publicado: (2026)
por: Scheinert, Dominik, et al.
Publicado: (2026)
LA-IMR: Latency-Aware, Predictive In-Memory Routing and Proactive Autoscaling for Tail-Latency-Sensitive Cloud Robotics
por: Seo, Eunil, et al.
Publicado: (2025)
por: Seo, Eunil, et al.
Publicado: (2025)
ORACL: Optimized Reasoning for Autoscaling via Chain of Thought with LLMs for Microservices
por: Bai, Haoyu, et al.
Publicado: (2026)
por: Bai, Haoyu, et al.
Publicado: (2026)
Metric Criticality Identification for Cloud Microservices
por: Singal, Akanksha, et al.
Publicado: (2025)
por: Singal, Akanksha, et al.
Publicado: (2025)
MSARS: A Meta-Learning and Reinforcement Learning Framework for SLO Resource Allocation and Adaptive Scaling for Microservices
por: Hu, Kan, et al.
Publicado: (2024)
por: Hu, Kan, et al.
Publicado: (2024)
SLICE: SLO-Driven Scheduling for LLM Inference on Edge Computing Devices
por: Chow, Will
Publicado: (2025)
por: Chow, Will
Publicado: (2025)
Joint Temporal-Structural Representation Learning for Distributed Fault Discrimination in Microservice Architectures
por: Xue, Yihan, et al.
Publicado: (2026)
por: Xue, Yihan, et al.
Publicado: (2026)
Areon: Latency-Friendly and Resilient Multi-Proposer Consensus
por: Castro-Castilla, Álvaro, et al.
Publicado: (2025)
por: Castro-Castilla, Álvaro, et al.
Publicado: (2025)
Contention-Aware Microservice Deployment in Collaborative Mobile Edge Networks
por: Ge, Xinlei, et al.
Publicado: (2024)
por: Ge, Xinlei, et al.
Publicado: (2024)
Shared Memory-Aware Latency-Sensitive Message Aggregation for Fine-Grained Communication
por: Chandrasekar, Kavitha, et al.
Publicado: (2024)
por: Chandrasekar, Kavitha, et al.
Publicado: (2024)
Low-Latency Layer-Aware Proactive and Passive Container Migration in Meta Computing
por: Liu, Mengjie, et al.
Publicado: (2024)
por: Liu, Mengjie, et al.
Publicado: (2024)
Cortex: Achieving Low-Latency, Cost-Efficient Remote Data Access For LLM via Semantic-Aware Knowledge Caching
por: Ruan, Chaoyi, et al.
Publicado: (2025)
por: Ruan, Chaoyi, et al.
Publicado: (2025)
Signalling Health for Improved Kubernetes Microservice Availability
por: Roberts, Jacob, et al.
Publicado: (2025)
por: Roberts, Jacob, et al.
Publicado: (2025)
Self-adaptive, Requirements-driven Autoscaling of Microservices
por: Nunes, João Paulo Karol Santos, et al.
Publicado: (2024)
por: Nunes, João Paulo Karol Santos, et al.
Publicado: (2024)
NotNets: Accelerating Microservices by Bypassing the Network
por: Alvaro, Peter, et al.
Publicado: (2024)
por: Alvaro, Peter, et al.
Publicado: (2024)
Chasing the Speed of Light: Low-Latency Planetary-Scale Adaptive Byzantine Consensus
por: Berger, Christian, et al.
Publicado: (2023)
por: Berger, Christian, et al.
Publicado: (2023)
TailBench++: Flexible Multi-Client, Multi-Server Benchmarking for Latency-Critical Workloads
por: Li, Zhilin, et al.
Publicado: (2025)
por: Li, Zhilin, et al.
Publicado: (2025)
Energy-aware Distributed Microservice Request Placement at the Edge
por: Toczé, Klervie, et al.
Publicado: (2024)
por: Toczé, Klervie, et al.
Publicado: (2024)
Energy Metrics for Edge Microservice Request Placement Strategies
por: Toczé, Klervie, et al.
Publicado: (2025)
por: Toczé, Klervie, et al.
Publicado: (2025)
EMLIO: Minimizing I/O Latency and Energy Consumption for Large-Scale AI Training
por: Jamil, Hasibul, et al.
Publicado: (2025)
por: Jamil, Hasibul, et al.
Publicado: (2025)
Overcoming Latency-bound Limitations of Distributed Graph Algorithms using the HPX Runtime System
por: Mohammadiporshokooh, Karame, et al.
Publicado: (2026)
por: Mohammadiporshokooh, Karame, et al.
Publicado: (2026)
Formal Specification for Fast ACS: Low-Latency File-Based Ordered Message Delivery at Scale
por: Gupta, Sushant Kumar, et al.
Publicado: (2025)
por: Gupta, Sushant Kumar, et al.
Publicado: (2025)
Flare: Leveraging Serverless Elasticity to Absorb Microservice Load Spikes
por: Dehigama, Dilina, et al.
Publicado: (2026)
por: Dehigama, Dilina, et al.
Publicado: (2026)
Service-Level Energy Modeling and Experimentation for Cloud-Native Microservices
por: Legler, Julian, et al.
Publicado: (2025)
por: Legler, Julian, et al.
Publicado: (2025)
ADApt: Edge Device Anomaly Detection and Microservice Replica Prediction
por: Mehran, Narges, et al.
Publicado: (2025)
por: Mehran, Narges, et al.
Publicado: (2025)
Auto-scaling Approaches for Microservice Applications: A Survey and Taxonomy
por: Xu, Minxian, et al.
Publicado: (2025)
por: Xu, Minxian, et al.
Publicado: (2025)
Artifact for Service-Level Energy Modeling and Experimentation for Cloud-Native Microservices
por: Legler, Julian
Publicado: (2026)
por: Legler, Julian
Publicado: (2026)
REACH: Reinforcement Learning for Adaptive Microservice Rescheduling in the Cloud-Edge Continuum
por: Bai, Xu, et al.
Publicado: (2025)
por: Bai, Xu, et al.
Publicado: (2025)
Mitigating Interference of Microservices with a Scoring Mechanism in Large-scale Clusters
por: Yang, Dingyu, et al.
Publicado: (2024)
por: Yang, Dingyu, et al.
Publicado: (2024)
Ejemplares similares
-
Reliable Microservice Tail Latency Prediction via Decoupled Dual-Stream Learning and Gradient Modulation
por: Qian, Wenzhuo, et al.
Publicado: (2025) -
HGraphScale: Hierarchical Graph Learning for Autoscaling Microservice Applications in Container-based Cloud Computing
por: Fang, Zhengxin, et al.
Publicado: (2025) -
StatuScale: Status-aware and Elastic Scaling Strategy for Microservice Applications
por: Wen, Linfeng, et al.
Publicado: (2024) -
Resilient Auto-Scaling of Microservice Architectures with Efficient Resource Management
por: Ahmad, Hussain, et al.
Publicado: (2025) -
Collaborative Evolution of Intelligent Agents in Large-Scale Microservice Systems
por: Li, Yilin, et al.
Publicado: (2025)