Large Language Models as Realistic Microservice Trace Generators
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Donghyun, Ravula, Sriram, Ha, Taemin, Dimakis, Alexandros G., Kim, Daehyeok, Akella, Aditya |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vulcan: Instance-Optimal Systems Heuristics Through LLM-Driven Search
by: Dwivedula, Rohit, et al.
Published: (2025)
by: Dwivedula, Rohit, et al.
Published: (2025)
Object as a Service: Simplifying Cloud-Native Development through Serverless Object Abstraction
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2024)
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2024)
Interferences within a certifiable design methodology for high-performance multi-core platforms
by: Khelassi, Mohamed Amine, et al.
Published: (2026)
by: Khelassi, Mohamed Amine, et al.
Published: (2026)
Who is in Charge here? Understanding How Runtime Configuration Affects Software along with Variables&Constants
by: Luo, Chaopeng, et al.
Published: (2025)
by: Luo, Chaopeng, et al.
Published: (2025)
Man-Made Heuristics Are Dead. Long Live Code Generators!
by: Dwivedula, Rohit, et al.
Published: (2025)
by: Dwivedula, Rohit, et al.
Published: (2025)
Fixed-Priority and EDF Schedules for ROS2 Graphs on Uniprocessor
by: Bell, Oren, et al.
Published: (2025)
by: Bell, Oren, et al.
Published: (2025)
Blink: CPU-Free LLM Inference by Delegating the Serving Stack to GPU and SmartNIC
by: Siavashi, Mohammad, et al.
Published: (2026)
by: Siavashi, Mohammad, et al.
Published: (2026)
Extending Data Spatial Semantics for Scale Agnostic Programming
by: Mars, Jason
Published: (2025)
by: Mars, Jason
Published: (2025)
Patchwork: A Unified Framework for RAG Serving
by: Hu, Bodun, et al.
Published: (2025)
by: Hu, Bodun, et al.
Published: (2025)
SoK: Microservice Architectures from a Dependability Perspective
by: Kažemaks, Dāvis, et al.
Published: (2025)
by: Kažemaks, Dāvis, et al.
Published: (2025)
Complexity at Scale: A Quantitative Analysis of an Alibaba Microservice Deployment
by: Winchester, Giles, et al.
Published: (2025)
by: Winchester, Giles, et al.
Published: (2025)
AdaptiFlow: An Extensible Framework for Event-Driven Autonomy in Cloud Microservices
by: Ndadji, Brice Arléon Zemtsop, et al.
Published: (2025)
by: Ndadji, Brice Arléon Zemtsop, et al.
Published: (2025)
Microservices-based Software Systems Reengineering: State-of-the-Art and Future Directions
by: Mohottige, Thakshila Imiya, et al.
Published: (2024)
by: Mohottige, Thakshila Imiya, et al.
Published: (2024)
Rethinking Inter-Process Communication with Memory Operation Offloading
by: Park, Misun, et al.
Published: (2026)
by: Park, Misun, et al.
Published: (2026)
Moving From Monolithic To Microservices Architecture for Multi-Agent Systems
by: Goyal, Muskaan, et al.
Published: (2025)
by: Goyal, Muskaan, et al.
Published: (2025)
TraceMesh: Scalable and Streaming Sampling for Distributed Traces
by: Chen, Zhuangbin, et al.
Published: (2024)
by: Chen, Zhuangbin, et al.
Published: (2024)
STaleX: A Spatiotemporal-Aware Adaptive Auto-scaling Framework for Microservices
by: Dashtbani, Majid, et al.
Published: (2025)
by: Dashtbani, Majid, et al.
Published: (2025)
FALCON: Pinpointing and Mitigating Stragglers for Large-Scale Hybrid-Parallel Training
by: Wu, Tianyuan, et al.
Published: (2024)
by: Wu, Tianyuan, et al.
Published: (2024)
BLITZSCALE: Fast and Live Large Model Autoscaling with O(1) Host Caching
by: Zhang, Dingyan, et al.
Published: (2024)
by: Zhang, Dingyan, et al.
Published: (2024)
OMEGA: A Low-Latency GNN Serving System for Large Graphs
by: Kim, Geon-Woo, et al.
Published: (2025)
by: Kim, Geon-Woo, et al.
Published: (2025)
CvxCluster: Solving Large, Complex, Granular Resource Allocation Problems 100-1000x Faster
by: Nnorom Jr, Obi, et al.
Published: (2026)
by: Nnorom Jr, Obi, et al.
Published: (2026)
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
by: Liu, Shu, et al.
Published: (2026)
by: Liu, Shu, et al.
Published: (2026)
Do Large Language Models Understand Performance Optimization?
by: Cui, Bowen, et al.
Published: (2025)
by: Cui, Bowen, et al.
Published: (2025)
Evaluating Asynchronous Semantics in Trace-Discovered Resilience Models: A Case Study on the OpenTelemetry Demo
by: Krasnovsky, Anatoly A.
Published: (2025)
by: Krasnovsky, Anatoly A.
Published: (2025)
Network Centrality as a New Perspective on Microservice Architecture
by: Bakhtin, Alexander, et al.
Published: (2025)
by: Bakhtin, Alexander, et al.
Published: (2025)
Mitigating context switching in densely packed Linux clusters with Latency-Aware Group Scheduling
by: Isstaif, Al Amjad Tawfiq, et al.
Published: (2025)
by: Isstaif, Al Amjad Tawfiq, et al.
Published: (2025)
DPC: A Distributed Page Cache over CXL
by: Bergman, Shai, et al.
Published: (2026)
by: Bergman, Shai, et al.
Published: (2026)
EdgeFlow: Fast Cold Starts for LLMs on Mobile Devices
by: Yan, Yongsheng, et al.
Published: (2026)
by: Yan, Yongsheng, et al.
Published: (2026)
Unlocking True Elasticity for the Cloud-Native Era with Dandelion
by: Kuchler, Tom, et al.
Published: (2025)
by: Kuchler, Tom, et al.
Published: (2025)
A Periodic Space of Distributed Computing: Vision & Framework
by: Salehi, Mohsen Amini, et al.
Published: (2026)
by: Salehi, Mohsen Amini, et al.
Published: (2026)
THEMIS: Time, Heterogeneity, and Energy Minded Scheduling for Fair Multi-Tenant Use in FPGAs
by: Karabulut, Emre, et al.
Published: (2024)
by: Karabulut, Emre, et al.
Published: (2024)
Formal Definitions and Performance Comparison of Consistency Models for Parallel File Systems
by: Wang, Chen, et al.
Published: (2024)
by: Wang, Chen, et al.
Published: (2024)
Mewz: Lightweight Execution Environment for WebAssembly with High Isolation and Portability using Unikernels
by: Ueda, Soichiro, et al.
Published: (2024)
by: Ueda, Soichiro, et al.
Published: (2024)
Taming Serverless Cold Starts Through OS Co-Design
by: Holmes, Ben, et al.
Published: (2025)
by: Holmes, Ben, et al.
Published: (2025)
Fix: externalizing network I/O in serverless computing
by: Deng, Yuhan, et al.
Published: (2025)
by: Deng, Yuhan, et al.
Published: (2025)
RAGDoll: Efficient Offloading-based Online RAG System on a Single GPU
by: Yu, Weiping, et al.
Published: (2025)
by: Yu, Weiping, et al.
Published: (2025)
Optimizing Task Scheduling in Heterogeneous Computing Environments: A Comparative Analysis of CPU, GPU, and ASIC Platforms Using E2C Simulator
by: Mohammadjafari, Ali, et al.
Published: (2024)
by: Mohammadjafari, Ali, et al.
Published: (2024)
Performance Isolation and Semantic Determinism in Efficient GPU Spatial Sharing
by: Yang, Zhenyuan, et al.
Published: (2026)
by: Yang, Zhenyuan, et al.
Published: (2026)
Telepathic Datacenters: Fast RPCs using Shared CXL Memory
by: Mahar, Suyash, et al.
Published: (2024)
by: Mahar, Suyash, et al.
Published: (2024)
Equilibria: Fair Multi-Tenant CXL Memory Tiering At Scale
by: Zhao, Kaiyang, et al.
Published: (2026)
by: Zhao, Kaiyang, et al.
Published: (2026)
Similar Items
-
Vulcan: Instance-Optimal Systems Heuristics Through LLM-Driven Search
by: Dwivedula, Rohit, et al.
Published: (2025) -
Object as a Service: Simplifying Cloud-Native Development through Serverless Object Abstraction
by: Lertpongrujikorn, Pawissanutt, et al.
Published: (2024) -
Interferences within a certifiable design methodology for high-performance multi-core platforms
by: Khelassi, Mohamed Amine, et al.
Published: (2026) -
Who is in Charge here? Understanding How Runtime Configuration Affects Software along with Variables&Constants
by: Luo, Chaopeng, et al.
Published: (2025) -
Man-Made Heuristics Are Dead. Long Live Code Generators!
by: Dwivedula, Rohit, et al.
Published: (2025)